Qwen Councils
arXiv is taking too long to respond. Please try again or narrow your search.
Showing downloaded papers while arXiv is unavailable.

Computer Science

arXiv preprints from January 1, 2026 through July 20, 2026 — 08:46:06 EST

0

Posted in cs.NI · 2026-01-21 · Francesco Rossato, Mattia Figaro, Alessandro Traspadini, Takayuki Shimizu, Chinmay Mahabal, Sanjeewa Herath, Chunghan Lee, Dogan Kutay Pekcan, Michele Zorzi, Marco Giordani

5G NR Non-Terrestrial Networks: Open Challenges for Full-Stack Protocol Design

As 5th generation (5G) networks continue to evolve, there is a growing interest toward the integration of Terrestrial Networks (TNs) and Non-Terrestrial Networks (NTNs). Specifically, NTNs leverage space/air base stations such as satellites, High Altitude Platforms (HAPs), and Unmanned Aerial Vehicles (UAVs) for expanding wireless...

💬 0 commentsarXiv:2601.14883v1PDF
0

Posted in cs.CV · 2026-01-21 · Zhe Chang, Haodong Jin, Ying Sun, Yan Song, Hui Yu

GAT-NeRF: Geometry-Aware-Transformer Enhanced Neural Radiance Fields for High-Fidelity 4D Facial Avatars

High-fidelity 4D dynamic facial avatar reconstruction from monocular video is a critical yet challenging task, driven by increasing demands for immersive virtual human applications. While Neural Radiance Fields (NeRF) have advanced scene representation, their capacity to capture high-frequency facial details, such as dynamic wrinkles...

💬 0 commentsarXiv:2601.14875v1PDF
0

Posted in cs.RO · 2026-01-21 · Yara Mahmoud, Yasheerah Yaqoot, Miguel Altamirano Cabrera, Dzmitry Tsetserukou

HumanoidVLM: Vision-Language-Guided Impedance Control for Contact-Rich Humanoid Manipulation

Humanoid robots must adapt their contact behavior to diverse objects and tasks, yet most controllers rely on fixed, hand-tuned impedance gains and gripper settings. This paper introduces HumanoidVLM, a vision-language driven retrieval framework that enables the Unitree G1 humanoid to select task-appropriate Cartesian impedance...

💬 0 commentsarXiv:2601.14874v1PDF
0

Posted in cs.RO · 2026-01-21 · Zejian Cui, Ferdinando Rodriguez y Baena

On-the-fly hand-eye calibration for the da Vinci surgical robot

In Robot-Assisted Minimally Invasive Surgery (RMIS), accurate tool localization is crucial to ensure patient safety and successful task execution. However, this remains challenging for cable-driven robots, such as the da Vinci robot, because erroneous encoder readings lead to pose estimation errors. In this study, we propose a...

💬 0 commentsarXiv:2601.14871v2PDF
0

Posted in cs.SE · 2026-01-21 · Martin Obaidi, Kushtrim Qengaj, Hannah Deters, Jakob Droste, Marc Herrmann, Kurt Schneider, Jil Klünder

Understanding Usefulness in Developer Explanations on Stack Overflow

Explanations are essential in software engineering (SE) and requirements communication, helping stakeholders clarify ambiguities, justify design choices, and build shared understanding. Online Q&A forums such as Stack Overflow provide large-scale settings where such explanations are produced and evaluated, offering valuable insights...

💬 0 commentsarXiv:2601.14865v1PDF
0

Posted in cs.LG · 2026-01-21 · Olaf Yunus Laitinen Imanov, Taner Yilmaz, Derya Umut Kulali

Strategic Doctrine Language Models (sdLM): A Learning-System Framework for Doctrinal Consistency and Geopolitical Forecasting

We introduce Strategic Doctrine Language Models (sdLM), a learning-system framework for multi-document strategic reasoning with doctrinal consistency constraints and calibrated uncertainty. The approach combines multi-document attention, temporal encoding, and a doctrine-consistency layer to improve long-horizon forecasting and plan...

💬 0 commentsarXiv:2601.14862v1PDF
0

Posted in cs.SE · 2026-01-21 · Tanja E. J. Vos, Tijs van der Storm, Alexander Serebrenik, Lionel Briand, Roberto Di Cosmo, J. -M Bruel, Benoît Combemale

Reclaiming Software Engineering as the Enabling Technology for the Digital Age

Software engineering is the invisible infrastructure of the digital age. Every breakthrough in artificial intelligence, quantum computing, photonics, and cybersecurity relies on advances in software engineering, yet the field is too often treated as a supportive digital component rather than as a strategic, enabling discipline. In...

💬 0 commentsarXiv:2601.14861v2PDF
0

Posted in cs.CL · 2026-01-21 · Motong Tian, Allen P. Wong, Mingjun Mao, Wangchunshu Zhou

HiNS: Hierarchical Negative Sampling for More Comprehensive Memory Retrieval Embedding Model

Memory-augmented language agents rely on embedding models for effective memory retrieval. However, existing training data construction overlooks a critical limitation: the hierarchical difficulty of negative samples and their natural distribution in human-agent interactions. In practice, some negatives are semantically close...

💬 0 commentsarXiv:2601.14857v1PDF
0

Posted in cs.LG · 2026-01-21 · Baojun Che, Yifan Chen, Daniel Zhengyu Huang, Xinying Mao, Weijie Wang

Adaptive Exponential Integration for Stable Gaussian Mixture Black-Box Variational Inference

Black-box variational inference (BBVI) with Gaussian mixture families offers a flexible approach for approximating complex posterior distributions without requiring gradients of the target density. However, standard numerical optimization methods often suffer from instability and inefficiency. We develop a stable and efficient...

💬 0 commentsarXiv:2601.14855v3PDF
0

Posted in cs.SD · 2026-01-21 · Viola Negroni, Luca Cuccovillo, Paolo Bestagini, Patrick Aichroth, Stefano Tubaro

Multi-Task Transformer for Explainable Speech Deepfake Detection via Formant Modeling

In this work, we introduce a multi-task transformer for speech deepfake detection, capable of predicting formant trajectories and voicing patterns over time, ultimately classifying speech as real or fake, and highlighting whether its decisions rely more on voiced or unvoiced regions. Building on a prior speaker-formant transformer...

💬 0 commentsarXiv:2601.14850v2PDF
0

Posted in cs.LG · 2026-01-21 · Mohamed Abouras, Catherine M. Elias

From Observation to Prediction: LSTM for Vehicle Lane Change Forecasting on Highway On/Off-Ramps

On and off-ramps are understudied road sections even though they introduce a higher level of variation in highway interactions. Predicting vehicles' behavior in these areas can decrease the impact of uncertainty and increase road safety. In this paper, the difference between this Area of Interest (AoI) and a straight highway section...

💬 0 commentsarXiv:2601.14848v1PDF
0

Posted in cs.LO · 2026-01-21 · Satoshi Kura, Marco Gaboardi, Taro Sekiyama, Hiroshi Unno

A Category-Theoretic Framework for Dependent Effect Systems

Graded monads refine traditional monads using effect annotations in order to describe quantitatively the computational effects that a program can generate. They have been successfully applied to a variety of formal systems for reasoning about effectful computations. However, existing categorical frameworks for graded monads do not...

💬 0 commentsarXiv:2601.14846v1PDF
0

Posted in cs.GR · 2026-01-21 · Zhe Chang, Haodong Jin, Yan Song, Hui Yu

CAG-Avatar: Cross-Attention Guided Gaussian Avatars for High-Fidelity Head Reconstruction

Creating high-fidelity, real-time drivable 3D head avatars is a core challenge in digital animation. While 3D Gaussian Splashing (3D-GS) offers unprecedented rendering speed and quality, current animation techniques often rely on a "one-size-fits-all" global tuning approach, where all Gaussian primitives are uniformly driven by a...

💬 0 commentsarXiv:2601.14844v1PDF
0

Posted in cs.CV · 2026-01-21 · Sidi Mohamed Sid El Moctar, Achraf Ait Laydi, Yousef El Mourabit, Hélène Bouvrais

MTFlow: Time-Conditioned Flow Matching for Microtubule Segmentation in Noisy Microscopy Images

Microtubules are cytoskeletal filaments that play essential roles in many cellular processes and are key therapeutic targets in several diseases. Accurate segmentation of microtubule networks is critical for studying their organization and dynamics but remains challenging due to filament curvature, dense crossings, and image noise. We...

💬 0 commentsarXiv:2601.14841v1PDF
0

Posted in cs.AI · 2026-01-21 · Abdelrhman Bassiouny, Tom Schierenbeck, Sorin Arion, Benjamin Alt, Naren Vasantakumaar, Giang Nguyen, Michael Beetz

Implementing Knowledge Representation and Reasoning with Object Oriented Design

This paper introduces KRROOD, a framework designed to bridge the integration gap between modern software engineering and Knowledge Representation & Reasoning (KR&R) systems. While Object-Oriented Programming (OOP) is the standard for developing complex applications, existing KR&R frameworks often rely on external ontologies and...

💬 0 commentsarXiv:2601.14840v1PDF
0

Posted in cs.RO · 2026-01-21 · B. Calmé, N. J. Greenidge, A. Metcalf, A. Bacchetti, G. Loza, D. Kpeglo, P. Lloyd, V. Pensabene, J. H. Chandler, P. Valdastri

Moving Beyond Compliance in Soft-Robotic Catheters Through Modularity for Precision Therapies

Soft robotic instruments could navigate delicate, tortuous anatomy more safely than rigid tools, but clinical adoption is limited by insufficient tip functionalization and real-time feedback at the tissue interface. Few sensing and therapeutic modules are compact, robust, and adaptable enough to measure, and respond to, subtle...

💬 0 commentsarXiv:2601.14837v1PDF
0

Posted in cs.AI · 2026-01-21 · Ben Schaper, Maxime Di Folco, Bernhard Kainz, Julia A. Schnabel, Cosmin I. Bercea

Measuring and Aligning Abstraction in Vision-Language Models with Medical Taxonomies

Vision-Language Models show strong zero-shot performance for chest X-ray classification, but standard flat metrics fail to distinguish between clinically minor and severe errors. This work investigates how to quantify and mitigate abstraction errors by leveraging medical taxonomies. We benchmark several state-of-the-art VLMs using...

💬 0 commentsarXiv:2601.14827v1PDF
0

Posted in cs.CL · 2026-01-21 · Yuxuan Cao, Zida Yang, Ye Wang

Comparative Study of Large Language Models on Chinese Film Script Continuation: An Empirical Analysis Based on GPT-5.2 and Qwen-Max

As large language models (LLMs) are increasingly applied to creative writing, their performance on culturally specific narrative tasks warrants systematic investigation. This study constructs the first Chinese film script continuation benchmark comprising 53 classic films, and designs a multi-dimensional evaluation framework comparing...

💬 0 commentsarXiv:2601.14826v1PDF
0

Posted in cs.LG · 2026-01-21 · Francisco Martínez, María P. Frías

Automated univariate time series forecasting with regression trees

This paper describes a methodology for automated univariate time series forecasting using regression trees and their ensembles: bagging and random forests. The key aspects that are addressed are: the use of an autoregressive approach and recursive forecasts, how to select the autoregressive features, how to deal with trending series...

💬 0 commentsarXiv:2602.00077v1PDF
0

Posted in cs.DL · 2026-01-21 · Martin Critelli

Archives, archival bond, and digital representation: A case study with the International Image Interoperability Framework

Within the archival sector, digitization has long been a strategic initiative to ensure greater availability of historical documents. In recent years, the promotion of guidelines and standards, combined with technological advancements, has established methodologies and best practices and developed tools to facilitate massive...

💬 0 commentsarXiv:2601.14823v1PDF
0

Posted in cs.CV · 2026-01-21 · Volodymyr Sydorskyi, Igor Krashenyi, Oleksii Yakubenko

Multimodal system for skin cancer detection

Melanoma detection is vital for early diagnosis and effective treatment. While deep learning models on dermoscopic images have shown promise, they require specialized equipment, limiting their use in broader clinical settings. This study introduces a multi-modal melanoma detection system using conventional photo images, making it more...

💬 0 commentsarXiv:2601.14822v2PDF
0

Posted in cs.SD · 2026-01-21 · Thomas Serre, Mathieu Fontaine, Éric Benhaim, Slim Essid

Contrastive Knowledge Distillation for Embedding Refinement in Personalized Speech Enhancement

Personalized speech enhancement (PSE) has shown convincing results when it comes to extracting a known target voice among interfering ones. The corresponding systems usually incorporate a representation of the target voice within the enhancement system, which is extracted from an enrollment clip of the target voice with upstream...

💬 0 commentsarXiv:2601.16235v1PDF
0

Posted in cs.CV · 2026-01-21 · Bert Ramlot, Martijn Courteaux, Peter Lambert, Glenn Van Wallendael

POTR: Post-Training 3DGS Compression

3D Gaussian Splatting (3DGS) has recently emerged as a promising contender to Neural Radiance Fields (NeRF) in 3D scene reconstruction and real-time novel view synthesis. 3DGS outperforms NeRF in training and inference speed but has substantially higher storage requirements. To remedy this downside, we propose POTR, a post-training...

💬 0 commentsarXiv:2601.14821v1PDF
0

Posted in cs.LG · 2026-01-21 · Christian Fiedler

Statistical Learning Theory for Distributional Classification

In supervised learning with distributional inputs in the two-stage sampling setup, relevant to applications like learning-based medical screening or causal learning, the inputs (which are probability distributions) are not accessible in the learning phase, but only samples thereof. This problem is particularly amenable to kernel-based...

💬 0 commentsarXiv:2601.14818v1PDF
0

Posted in cs.CY · 2026-01-21 · Pierre Schaus, Guillaume Derval, Augustin Delecluse

ICLF: An Immersive Code Learning Framework based on Git for Teaching and Evaluating Student Programming Projects

Programming projects are essential in computer science education for bridging theory with practice and introducing students to tools like Git, IDEs, and debuggers. However, designing and evaluating these projects (especially in MOOCs)can be challenging. We propose the Immersive Code Learning Framework (ICLF), a scalable Git-based...

💬 0 commentsarXiv:2601.14814v1PDF