Qwen Councils

Computer Science

arXiv preprints from January 1, 2026 through July 20, 2026 — 03:40:02 EST

0

Posted in cs.SE · 2026-01-19 · Xue Jiang, Ge Li, Jiaru Qian, Xianjie Shi, Chenjie Li, Hao Zhu, Ziyu Wang, Jielun Zhang, Zheyu Zhao, Lingwei Wu, Kechi Zhang, Jia Li, Wenpin Jiao, Zhi Jin, Yihong Dong

KOCO-BENCH: Can Large Language Models Leverage Domain Knowledge in Software Development?

Large language models (LLMs) excel at general programming but struggle with domain-specific software development, necessitating domain specialization methods for LLMs to learn and utilize domain knowledge and data. However, existing domain-specific code benchmarks cannot evaluate the effectiveness of domain specialization methods,...

💬 0 commentsarXiv:2601.13240v3PDF
0

Posted in cs.CV · 2026-01-19 · Chengyin Hu, Xiang Chen, Zhe Jia, Weiwen Shi, Fengyu Zhang, Jiujiang Guo, Yiwei Wei

A Semantic Decoupling-Based Two-Stage Rainy-Day Attack for Revealing Weather Robustness Deficiencies in Vision-Language Models

Vision-Language Models (VLMs) are trained on image-text pairs collected under canonical visual conditions and achieve strong performance on multimodal tasks. However, their robustness to real-world weather conditions, and the stability of cross-modal semantic alignment under such structured perturbations, remain insufficiently...

💬 0 commentsarXiv:2601.13238v1PDF
0

Posted in cs.HC · 2026-01-19 · Drishti Goel, Jeongah Lee, Qiuyue Joy Zhong, Violeta J. Rodriguez, Daniel S. Brown, Ravi Karkar, Dong Whi Yoo, Koustuv Saha

RubRIX: Rubric-Driven Risk Mitigation in Caregiver-AI Interactions

Caregivers seeking AI-mediated support express complex needs -- information-seeking, emotional validation, and distress cues -- that warrant careful evaluation of response safety and appropriateness. Existing AI evaluation frameworks, primarily focused on general risks (toxicity, hallucinations, policy violations, etc), may not...

💬 0 commentsarXiv:2601.13235v1PDF
0

Posted in cs.CV · 2026-01-19 · Md. Nishan Khan, Kazi Shahriar Sanjid, Md. Tanzim Hossain, Asib Mostakim Fony, Istiak Ahmed, M. Monir Uddin

ConvMambaNet: A Hybrid CNN-Mamba State Space Architecture for Accurate and Real-Time EEG Seizure Detection

Epilepsy is a chronic neurological disorder marked by recurrent seizures that can severely impact quality of life. Electroencephalography (EEG) remains the primary tool for monitoring neural activity and detecting seizures, yet automated analysis remains challenging due to the temporal complexity of EEG signals. This study introduces...

💬 0 commentsarXiv:2601.13234v1PDF
0

Posted in cs.AI · 2026-01-19 · Bolin Chen, Dex Doksoo Lee, Wei "Wayne'' Chen, Wei Chen

RAG: A Random-Forest-Based Generative Design Framework for Uncertainty-Aware Design of Metamaterials with Complex Functional Response Requirements

Metamaterials design for advanced functionality often entails the inverse design on nonlinear and condition-dependent responses (e.g., stress-strain relation and dispersion relation), which are described by continuous functions. Most existing design methods focus on vector-valued responses (e.g., Young's modulus and bandgap width),...

💬 0 commentsarXiv:2601.13233v1PDF
0

Posted in cs.RO · 2026-01-19 · Kourosh Darvish, Arjun Sohal, Abhijoy Mandal, Hatem Fakhruldeen, Nikola Radulov, Zhengxue Zhou, Satheeshkumar Veeramani, Joshua Choi, Sijie Han, Brayden Zhang, Jeeyeoun Chae, Alex Wright, Yijie Wang, Hossein Darvish, Yuchi Zhao, Gary Tom, Han Hao, Miroslav Bogdanovic, Gabriella Pizzuto, Andrew I. Cooper, Alán Aspuru-Guzik, Florian Shkurti, Animesh Garg

MATTERIX: toward a digital twin for robotics-assisted chemistry laboratory automation

Accelerated materials discovery is critical for addressing global challenges. However, developing new laboratory workflows relies heavily on real-world experimental trials, and this can hinder scalability because of the need for numerous physical make-and-test iterations. Here we present MATTERIX, a multiscale, graphics processing...

💬 0 commentsarXiv:2601.13232v1PDF
0

Posted in cs.CL · 2026-01-19 · Tianqi Du, Lizhe Fang, Weijie Yang, Chenheng Zhang, Zeming Wei, Yifei Wang, Yisen Wang

Autoregressive Models Rival Diffusion Models at ANY-ORDER Generation

Diffusion language models enable any-order generation and bidirectional conditioning, offering appealing flexibility for tasks such as infilling, rewriting, and self-correction. However, their formulation-predicting one part of a sequence from another within a single-step dependency-limits modeling depth and often yields lower sample...

💬 0 commentsarXiv:2601.13228v1PDF
0

Posted in cs.IR · 2026-01-19 · Laura Dietz, Bryan Li, Eugene Yang, Dawn Lawrie, William Walden, James Mayfield

Insider Knowledge: How Much Can RAG Systems Gain from Evaluation Secrets?

RAG systems are increasingly evaluated and optimized using LLM judges, an approach that is rapidly becoming the dominant paradigm for system assessment. Nugget-based approaches in particular are now embedded not only in evaluation frameworks but also in the architectures of RAG systems themselves. While this integration can lead to...

💬 0 commentsarXiv:2601.13227v2PDF
0

Posted in cs.CV · 2026-01-19 · Tim Lachmann, Alexandra Israelsson, Christina Tornberg, Teimuraz Saghinadze, Michal Balazia, Philipp Müller, Petri Laukka

Not all Blends are Equal: The BLEMORE Dataset of Blended Emotion Expressions with Relative Salience Annotations

Humans often experience not just a single basic emotion at a time, but rather a blend of several emotions with varying salience. Despite the importance of such blended emotions, most video-based emotion recognition approaches are designed to recognize single emotions only. The few approaches that have attempted to recognize blended...

💬 0 commentsarXiv:2601.13225v1PDF
0

Posted in cs.PL · 2026-01-19 · Michael Hanus, Steven Libby

Functional Logic Program Transformations

Many tools used to process programs, like compilers, analyzers, or verifiers, perform transformations on their intermediate program representation, like abstract syntax trees. Implementing such program transformations is a non-trivial task, since it is necessary to iterate over the complete syntax tree and apply various...

💬 0 commentsarXiv:2601.13224v1PDF
0

Posted in cs.IR · 2026-01-19 · Laura Dietz, Bryan Li, Gabrielle Liu, Jia-Huei Ju, Eugene Yang, Dawn Lawrie, William Walden, James Mayfield

Incorporating Q&A Nuggets into Retrieval-Augmented Generation

RAGE systems integrate ideas from automatic evaluation (E) into Retrieval-augmented Generation (RAG). As one such example, we present Crucible, a Nugget-Augmented Generation System that preserves explicit citation provenance by constructing a bank of Q&A nuggets from retrieved documents and uses them to guide extraction, selection,...

💬 0 commentsarXiv:2601.13222v2PDF
0

Posted in cs.DS · 2026-01-19 · Paolo Ferragina, Francesco Tosoni

The Energy-Throughput Trade-off in Lossless-Compressed Source Code Storage

Retrieving data from large-scale source code archives is vital for AI training, neural-based software analysis, and information retrieval, to cite a few. This paper studies and experiments with the design of a compressed key-value store for the indexing of large-scale source code datasets, evaluating its trade-off among three primary...

💬 0 commentsarXiv:2601.13220v1PDF
0

Posted in cs.CV · 2026-01-19 · Igor Vozniak, Philipp Mueller, Nils Lipp, Janis Sprenger, Konstantin Poddubnyy, Davit Hovhannisyan, Christian Mueller, Andreas Bulling, Philipp Slusallek

ObjectVisA-120: Object-based Visual Attention Prediction in Interactive Street-crossing Environments

The object-based nature of human visual attention is well-known in cognitive science, but has only played a minor role in computational visual attention models so far. This is mainly due to a lack of suitable datasets and evaluation metrics for object-based attention. To address these limitations, we present ObjectVisA-120 -- a novel...

💬 0 commentsarXiv:2601.13218v2PDF
0

Posted in cs.CL · 2026-01-19 · Bingsen Chen, Boyan Li, Ping Nie, Yuyu Zhang, Xi Ye, Chen Zhao

Beyond Single-shot Writing: Deep Research Agents are Unreliable at Multi-turn Report Revision

Existing benchmarks for Deep Research Agents (DRAs) treat report generation as a single-shot writing task, which fundamentally diverges from how human researchers iteratively draft and revise reports via self-reflection or peer feedback. Whether DRAs can reliably revise reports with user feedback remains unexplored. We introduce Mr...

💬 0 commentsarXiv:2601.13217v1PDF
0

Posted in cs.IT · 2026-01-19 · Ataher Sams, Besma Smida

On the Reliability of Estimation Bounds in Low-SNR Bistatic ISAC

This paper explores a bistatic Integrated Sensing and Communication (ISAC) framework, where a base station transmits communication signal that serve both direct communication with a user and multi-target parameter estimation through reflections captured by a separate sensing receiver. We assume that the instantaneous knowledge of the...

💬 0 commentsarXiv:2601.13216v1PDF
0

Posted in cs.IT · 2026-01-19 · Zheyu Wu, Junjie Ma, Ya-Feng Liu, Bruno Clerckx

An AMP-Based Asymptotic Analysis For Nonlinear One-Bit Precoding

This paper focuses on the asymptotic analysis of a class of nonlinear one-bit precoding schemes under Rayleigh fading channels. The considered scheme employs a convex-relaxation-then-quantization (CRQ) approach to the well-known minimum mean square error (MMSE) model, which includes the classical one-bit precoder SQUID as a special...

💬 0 commentsarXiv:2601.13214v1PDF
0

Posted in cs.NI · 2026-01-19 · Joao F. Santos, Arshia Zolghadr, Scott Kuzdeba, Jacek Kibiłda

Conflict Detection in AI-RAN: Efficient Interaction Learning and Autonomous Graph Reconstruction

Artificial Intelligence (AI)-native mobile networks represent a fundamental step toward 6G, where learning, inference, and decision making are embedded into the Radio Access Network (RAN) itself. In such networks, multiple AI agents optimize the network to achieve distinct and often competing objectives. As such, conflicts become...

💬 0 commentsarXiv:2601.13213v2PDF
0

Posted in cs.CV · 2026-01-19 · Vikram R Lakkavalli

Rethinking Skip Connections: Additive U-Net for Robust and Interpretable Denoising

Skip connections are central to U-Net architectures for image denoising, but standard concatenation doubles channel dimensionality and obscures information flow, allowing uncontrolled noise transfer. We propose the Additive U-Net, which replaces concatenative skips with gated additive connections. Each skip pathway is scaled by a...

💬 0 commentsarXiv:2601.13208v1PDF
0

Posted in cs.CV · 2026-01-19 · Jinnao Li, Zijian Chen, Tingzhu Chen, Changbo Wang

GTPred: Benchmarking MLLMs for Interpretable Geo-localization and Time-of-capture Prediction

Geo-localization aims to infer the geographic location where an image was captured using observable visual evidence. Traditional methods achieve impressive results through large-scale training on massive image corpora. With the emergence of multi-modal large language models (MLLMs), recent studies have explored their applications in...

💬 0 commentsarXiv:2601.13207v1PDF
0

Posted in cs.AI · 2026-01-19 · Neil K. R. Sehgal, Sharath Chandra Guntuku, Lyle Ungar

Real-Time Deadlines Reveal Temporal Awareness Failures in LLM Strategic Dialogues

Large Language Models (LLMs) generate text token-by-token in discrete time, yet real-world communication, from therapy sessions to business negotiations, critically depends on continuous time constraints. Current LLM architectures and evaluation protocols rarely test for temporal awareness under real-time deadlines. We use simulated...

💬 0 commentsarXiv:2601.13206v1PDF
0

Posted in cs.SD · 2026-01-19 · Yang Wang, Yiqi Liu, Chenghao Xiao, Chenghua Lin

The Achilles' Heel of Angular Margins: A Chebyshev Polynomial Fix for Speaker Verification

Angular margin losses, such as AAM-Softmax, have become the de facto in speaker and face verification. Their success hinges on directly manipulating the angle between features and class prototypes. However, this manipulation relies on the arccos function to recover the angle, introducing a significant yet overlooked source of training...

💬 0 commentsarXiv:2601.13198v1PDF
0

Posted in cs.CR · 2026-01-19 · Aravind B, Anirud R. S., Sai Surya Teja N, Bala Subrahmanya Sriranga Navaneeth A, Karthika R, Mohankumar N

Diffusion-Driven Synthetic Tabular Data Generation for Enhanced DoS/DDoS Attack Classification

Class imbalance refers to a situation where certain classes in a dataset have significantly fewer samples than oth- ers, leading to biased model performance. Class imbalance in network intrusion detection using Tabular Denoising Diffusion Probability Models (TabDDPM) for data augmentation is ad- dressed in this paper. Our approach...

💬 0 commentsarXiv:2601.13197v2PDF
0

Posted in cs.RO · 2026-01-19 · Jacob Swindell, Marija Popović, Riccardo Polvara

Active Informative Planning for UAV-based Weed Mapping using Discrete Gaussian Process Representations

Accurate agricultural weed mapping using unmanned aerial vehicles (UAVs) is crucial for precision farming. While traditional methods rely on rigid, pre-defined flight paths and intensive offline processing, informative path planning (IPP) offers a way to collect data adaptively where it is most needed. Gaussian process (GP) mapping...

💬 0 commentsarXiv:2601.13196v1PDF
0

Posted in cs.LG · 2026-01-19 · Vittoria De Pellegrini, Tariq Alkhalifah

LAViG-FLOW: Latent Autoregressive Video Generation for Fluid Flow Simulations

Modeling and forecasting subsurface multiphase fluid flow fields underpin applications ranging from geological CO2 sequestration (GCS) operations to geothermal production. This is essential for ensuring both operational performance and long-term safety. While high fidelity multiphase simulators are widely used for this purpose, they...

💬 0 commentsarXiv:2601.13190v2PDF
0

Posted in cs.HC · 2026-01-19 · Patrick Yung Kang Lee, Jessica Y. Bo, Zixin Zhao, Paula Akemi Aoyagui, Matthew Varona, Ashton Anderson, Anastasia Kuzminykh, Fanny Chevalier, Carolina Nobre

Large Language Lovers: Lived Experiences of Negotiating Agency and Platform Control in AI Companionship

Individuals are turning to increasingly anthropomorphic, general-purpose chatbots for AI companionship, rather than roleplay-specific platforms. However, not much is known about how individuals perceive and conduct their relationships with general-purpose chatbots. We triangulated community discussions on Reddit (41k+ posts and...

💬 0 commentsarXiv:2601.13188v3PDF