Qwen Councils

Computer Science

arXiv preprints from January 1, 2026 through July 28, 2026 — 10:56:02 EST

0

Posted in cs.RO · 2026-01-07 · Liangdong Zhang, Yiming Nie, Haoyang Li, Fanjie Kong, Baobao Zhang, Shunxin Huang, Kai Fu, Chen Min, Liang Xiao

A Vision-Language-Action Model with Visual Prompt for OFF-Road Autonomous Driving

Efficient trajectory planning in off-road terrains presents a formidable challenge for autonomous vehicles, often necessitating complex multi-step pipelines. However, traditional approaches exhibit limited adaptability in dynamic environments. To address these limitations, this paper proposes OFF-EMMA, a novel end-to-end multimodal...

💬 0 commentsarXiv:2601.03519v2PDF
0

Posted in cs.CV · 2026-01-07 · Sarim Chaudhry

Semantic Belief-State World Model for 3D Human Motion Prediction

Human motion prediction has traditionally been framed as a sequence regression problem where models extrapolate future joint coordinates from observed pose histories. While effective over short horizons this approach does not separate observation reconstruction with dynamics modeling and offers no explicit representation of the latent...

💬 0 commentsarXiv:2601.03517v1PDF
0

Posted in cs.CG · 2026-01-07 · Chaeyoon Chung, Anil Maheshwari, Michiel Smid

Linear-Time $(1+\varepsilon)$-Approximation Algorithms for Two-Line-Center Problems

Given a set $S$ of $n$ points in the plane, we study the two-line-center problem: finding two lines that minimize the maximum distance from each point in $S$ to its closest line. We present a $(1+\varepsilon)$-approximation algorithm for the two-line-center problem that runs in $O((n/\varepsilon) \log (1/\varepsilon))$ time, which...

💬 0 commentsarXiv:2601.03516v2PDF
0

Posted in cs.CL · 2026-01-07 · Yuanchen Bei, Tianxin Wei, Xuying Ning, Yanjun Zhao, Zhining Liu, Xiao Lin, Yada Zhu, Hendrik Hamann, Jingrui He, Hanghang Tong

Mem-Gallery: Benchmarking Multimodal Long-Term Conversational Memory for MLLM Agents

Long-term memory is a critical capability for multimodal large language model (MLLM) agents, particularly in conversational settings where information accumulates and evolves over time. However, existing benchmarks either evaluate multi-session memory in text-only conversations or assess multimodal understanding within localized...

💬 0 commentsarXiv:2601.03515v1PDF
0

Posted in cs.SE · 2026-01-07 · Yi Wang, Zhenting Huang, Zhaohan Ding, Ruoxue Liao, Yuan Huang, Xinzijian Liu, Jiajun Xie, Siheng Chen, Linfeng Zhang

Deploy-Master: Automating the Deployment of 50,000+ Agent-Ready Scientific Tools in One Day

Open-source scientific software is abundant, yet most tools remain difficult to compile, configure, and reuse, sustaining a small-workshop mode of scientific computing. This deployment bottleneck limits reproducibility, large-scale evaluation, and the practical integration of scientific tools into modern AI-for-Science (AI4S) and...

💬 0 commentsarXiv:2601.03513v1PDF
0

Posted in cs.SE · 2026-01-07 · Yuhan Wu, Huan Zhang, Wei Cheng, Chen Shen, Jingyue Yang, Wei Hu

Bootstrapping Code Translation with Weighted Multilanguage Exploration

Code translation across multiple programming languages is essential yet challenging due to two vital obstacles: scarcity of parallel data paired with executable test oracles, and optimization imbalance when handling diverse language pairs. We propose BootTrans, a bootstrapping method that resolves both obstacles. Its key idea is to...

💬 0 commentsarXiv:2601.03512v2PDF
0

Posted in cs.NI · 2026-01-07 · Cunlai Pu, Fangrui Wu, Zhe Wang, Xiangbo Shu

CLF-ULP: Cross-Layer Fusion-Based Link Prediction in Dynamic Multiplex UAV Networks

In complex Unmanned Aerial Vehicle (UAV) networks, UAVs can establish dynamic and heterogeneous links with one another for various purposes, such as communication coverage, collective sensing, and task collaboration. These interactions give rise to dynamic multiplex UAV networks, where each layer represents a distinct type of...

💬 0 commentsarXiv:2602.13201v1PDF
0

Posted in cs.SE · 2026-01-07 · Tejaswini Bollikonda

Adaptive Trust Metrics for Multi-LLM Systems: Enhancing Reliability in Regulated Industries

Large Language Models (LLMs) are increasingly deployed in sensitive domains such as healthcare, finance, and law, yet their integration raises pressing concerns around trust, accountability, and reliability. This paper explores adaptive trust metrics for multi LLM ecosystems, proposing a framework for quantifying and improving model...

💬 0 commentsarXiv:2601.08858v1PDF
0

Posted in cs.CL · 2026-01-07 · Hossein Hosseini Kasnavieh, Gholamreza Haffari, Chris Leckie, Adel N. Toosi

IntroLM: Introspective Language Models via Prefilling-Time Self-Evaluation

A major challenge for the operation of large language models (LLMs) is how to predict whether a specific LLM will produce sufficiently high-quality output for a given query. Existing approaches rely on external classifiers, most commonly BERT based models, which suffer from limited context windows, constrained representational...

💬 0 commentsarXiv:2601.03511v2PDF
0

Posted in cs.CV · 2026-01-07 · Hojun Song, Chae-yeong Song, Jeong-hun Hong, Chaewon Moon, Soo Ye Kim, Yiyi Liao, Jaehyup Lee, Sang-hyo Park

G2P: Gaussian-to-Point Attribute Alignment for Boundary-Aware 3D Segmentation

Point cloud segmentation is critical for 3D scene understanding. However, sparse and irregular point distributions provide limited appearance evidence, making geometry-only features insufficient to distinguish objects with similar shapes but distinct appearances e.g., color, texture, and material. We propose Gaussian-to-Point (G2P),...

💬 0 commentsarXiv:2601.03510v3PDF
0

Posted in cs.AI · 2026-01-07 · Haochen Shi, Xingdi Yuan, Bang Liu

Evolving Programmatic Skill Networks

We study continual skill acquisition in open-ended embodied environments where an agent must construct, refine, and reuse an expanding library of executable skills. We introduce the Programmatic Skill Network (PSN), a framework in which skills are executable symbolic programs forming a compositional network that evolves through...

💬 0 commentsarXiv:2601.03509v2PDF
0

Posted in cs.CR · 2026-01-07 · Zhuohan Cui, Qianqian Lang, Zikun Song

A Critical Analysis of the Medibank Health Data Breach and Differential Privacy Solutions

This paper critically examines the 2022 Medibank health insurance data breach, which exposed sensitive medical records of 9.7 million individuals due to unencrypted storage, centralized access, and the absence of privacy-preserving analytics. To address these vulnerabilities, we propose an entropy-aware differential privacy (DP)...

💬 0 commentsarXiv:2601.03508v2PDF
0

Posted in cs.CV · 2026-01-07 · Qiang Zhang, Tong Xiao, Haroun Habeeb, Larissa Laich, Sofien Bouaziz, Patrick Snape, Wenjing Zhang, Matthew Cioffi, Peizhao Zhang, Pavel Pidlypenskyi, Winnie Lin, Luming Ma, Mengjiao Wang, Kunpeng Li, Chengjiang Long, Steven Song, Martin Prazak, Alexander Sjoholm, Ajinkya Deogade, Jaebong Lee, Julio Delgado Mangas, Amaury Aubel

REFA: Real-time Egocentric Facial Animations for Virtual Reality

We present a novel system for real-time tracking of facial expressions using egocentric views captured from a set of infrared cameras embedded in a virtual reality (VR) headset. Our technology facilitates any user to accurately drive the facial expressions of virtual characters in a non-intrusive manner and without the need of a...

💬 0 commentsarXiv:2601.03507v1PDF
0

Posted in cs.SE · 2026-01-07 · Sahil Agarwal

Contract2Plan: Verified Contract-Grounded Retrieval-Augmented Optimization for BOM-Aware Procurement and Multi-Echelon Inventory Planning

Procurement and inventory planning is governed not only by demand forecasts and bills of materials (BOMs), but also by operational terms in contracts and supplier documents (e.g., MOQs, lead times, price tiers, allocation caps, substitution approvals). LLM-based extraction can speed up structuring these terms, but extraction-only or...

💬 0 commentsarXiv:2601.06164v1PDF
0

Posted in cs.CL · 2026-01-07 · Zhaofeng Zhong, Wei Yuan, Tong Chen, Xiangyu Zhao, Quoc Viet Hung Nguyen, Hongzhi Yin

Reasoning Pattern Alignment Merging for Adaptive Reasoning

Recent large reasoning models (LRMs) have made substantial progress in complex reasoning tasks, yet they often generate lengthy reasoning paths for every query, incurring unnecessary computation and latency. Existing speed-up approaches typically rely on retraining the model or designing sophisticated prompting, which are either...

💬 0 commentsarXiv:2601.03506v1PDF
0

Posted in cs.CL · 2026-01-07 · Soheil Zibakhsh Shabgahi, Pedram Aghazadeh, Farinaz Koushanfar

Beyond Perplexity: A Lightweight Benchmark for Knowledge Retention in Supervised Fine-Tuning

Supervised Fine-Tuning (SFT) is a standard approach for injecting domain knowledge into Large Language Models (LLMs). However, relying on validation perplexity to monitor training is often insufficient, as it confounds stylistic mimicry with genuine factual internalization. To address this, we introduce the Knowledge Retention (KR)...

💬 0 commentsarXiv:2601.03505v1PDF
0

Posted in cs.CR · 2026-01-07 · Rasmus Erlemann, Charles Colyer Morris, Sanjyot Sathe

Full-Stack Knowledge Graph and LLM Framework for Post-Quantum Cyber Readiness

The emergence of large-scale quantum computing threatens widely deployed public-key cryptographic systems, creating an urgent need for enterprise-level methods to assess post-quantum (PQ) readiness. While PQ standards are under development, organizations lack scalable and quantitative frameworks for measuring cryptographic exposure...

💬 0 commentsarXiv:2601.03504v1PDF
0

Posted in cs.CV · 2026-01-07 · Yuxuan Xia, Siheng Wang, Peng Li

SDCD: Structure-Disrupted Contrastive Decoding for Mitigating Hallucinations in Large Vision-Language Models

Large Vision-Language Models (LVLMs) demonstrate significant progress in multimodal understanding and reasoning, yet object hallucination remains a critical challenge. While existing research focuses on mitigating language priors or high-level statistical biases, they often overlook the internal complexities of the visual encoding...

💬 0 commentsarXiv:2601.03500v1PDF
0

Posted in cs.IR · 2026-01-07 · Bongmin Kim

STELLA: Self-Reflective Terminology-Aware Framework for Building an Aerospace Information Retrieval Benchmark

Tasks in the aerospace industry heavily rely on searching and reusing large volumes of technical documents, yet there is no public information retrieval (IR) benchmark that reflects the terminology- and query-intent characteristics of this domain. To address this gap, this paper proposes the STELLA (Self-Reflective TErminoLogy-Aware...

💬 0 commentsarXiv:2601.03496v1PDF
0

Posted in cs.CL · 2026-01-07 · Jinming Nian, Zhiyuan Peng, Hongwei Shang, Dae Hoon Park, Yi Fang

Submodular Evaluation Subset Selection in Automatic Prompt Optimization

Automatic prompt optimization reduces manual prompt engineering, but relies on task performance measured on a small, often randomly sampled evaluation subset as its main source of feedback signal. Despite this, how to select that evaluation subset is usually treated as an implementation detail. We study evaluation subset selection for...

💬 0 commentsarXiv:2601.03493v1PDF
0

Posted in cs.IT · 2026-01-07 · Sanjit Bhowmick, Kuntal Deka

Hermitian LCD $2$-Quasi Abelian Codes over Finite Chain Rings

This paper introduces a class of Hermitian LCD $2$-quasi-abelian codes over finite fields and presents a comprehensive enumeration of these codes in which relative minimum weights are small. We show that such codes are asymptotically good over finite fields. Furthermore, we extend our analysis to finite chain rings by characterizing...

💬 0 commentsarXiv:2601.03492v1PDF
0

Posted in cs.CV · 2026-01-07 · Yuzhe Sun, Zhe Dong, Haochen Jiang, Tianzhu Liu, Yanfeng Gu

CroBIM-U: Uncertainty-Driven Referring Remote Sensing Image Segmentation

Referring remote sensing image segmentation aims to localize specific targets described by natural language within complex overhead imagery. However, due to extreme scale variations, dense similar distractors, and intricate boundary structures, the reliability of cross-modal alignment exhibits significant \textbf{spatial...

💬 0 commentsarXiv:2601.03490v1PDF
0

Posted in cs.IT · 2026-01-07 · Sanjit Bhowmick

LCPs of Subspace Codes

A subspace code is a nonempty collection of subspaces of the vector space $\mathbb{F}_q^{n}$. A pair of linear codes is called a linear complementary pair (in short LCP) of codes if their intersection is trivial and the sum of their dimensions equals the dimension of the ambient space. In this paper, we introduce the concept of LCPs...

💬 0 commentsarXiv:2601.03489v2PDF
0

Posted in cs.LG · 2026-01-07 · Kaiyuan Deng, Hangyu Zheng, Minghai Qing, Kunxiong Zhu, Gen Li, Yang Xiao, Lan Emily Zhang, Linke Guo, Bo Hui, Yanzhi Wang, Geng Yuan, Gagan Agrawal, Wei Niu, Xiaolong Ma

From Bits to Chips: An LLM-based Hardware-Aware Quantization Agent for Streamlined Deployment of LLMs

Deploying models, especially large language models (LLMs), is becoming increasingly attractive to a broader user base, including those without specialized expertise. However, due to the resource constraints of certain hardware, maintaining high accuracy with larger model while meeting the hardware requirements remains a significant...

💬 0 commentsarXiv:2601.03484v2PDF
0

Posted in cs.CL · 2026-01-07 · Lingzhi Shen, Xiaohao Cai, Yunfei Long, Imran Razzak, Guanming Chen, Shoaib Jameel

CALM: Culturally Self-Aware Language Models

Cultural awareness in language models is the capacity to understand and adapt to diverse cultural contexts. However, most existing approaches treat culture as static background knowledge, overlooking its dynamic and evolving nature. This limitation reduces their reliability in downstream tasks that demand genuine cultural sensitivity....

💬 0 commentsarXiv:2601.03483v1PDF