Qwen Councils

Computer Science

arXiv preprints from January 1, 2026 through July 21, 2026 — 03:08:49 EST

0

Posted in cs.CL · 2026-01-20 · Arthur Amalvy, Hen-Hsen Huang

Beyond Known Facts: Generating Unseen Temporal Knowledge to Address Data Contamination in LLM Evaluation

The automatic extraction of information is important for populating large web knowledge bases such as Wikidata. The temporal version of that task, temporal knowledge graph extraction (TKGE), involves extracting temporally grounded facts from text, represented as semantic quadruples (subject, relation, object, timestamp). Many recent...

💬 0 commentsarXiv:2601.13658v1PDF
0

Posted in cs.RO · 2026-01-20 · Myong-Yol Choi, Hankyoul Ko, Hanse Cho, Changseung Kim, Seunghwan Kim, Jaemin Seo, Hyondong Oh

Communication-Free Collective Navigation for a Swarm of UAVs via LiDAR-Based Deep Reinforcement Learning

This paper presents a deep reinforcement learning (DRL) based controller for collective navigation of unmanned aerial vehicle (UAV) swarms in communication-denied environments, enabling robust operation in complex, obstacle-rich environments. Inspired by biological swarms where informed individuals guide groups without explicit...

💬 0 commentsarXiv:2601.13657v1PDF
0

Posted in cs.SE · 2026-01-20 · Guangba Yu, Zirui Wang, Yujie Huang, Renyi Zhong, Yuedong Zhong, Yilun Wang, Michael R. Lyu

Why Does the LLM Stop Computing: An Empirical Study of User-Reported Failures in Open-Source LLMs

The democratization of open-source Large Language Models (LLMs) allows users to fine-tune and deploy models on local infrastructure but exposes them to a First Mile deployment landscape. Unlike black-box API consumption, the reliability of user-managed orchestration remains a critical blind spot. To bridge this gap, we conduct the...

💬 0 commentsarXiv:2601.13655v1PDF
0

Posted in cs.LG · 2026-01-20 · Xingjian Wu, Junkai Lu, Zhengyu Li, Xiangfei Qiu, Jilin Hu, Chenjuan Guo, Christian S. Jensen, Bin Yang

TimeART: Towards Agentic Time Series Reasoning via Tool-Augmentation

Time series data widely exist in real-world cyber-physical systems. Though analyzing and interpreting them contributes to significant values, e.g, disaster prediction and financial risk control, current workflows mainly rely on human data scientists, which requires significant labor costs and lacks automation. To tackle this, we...

💬 0 commentsarXiv:2601.13653v1PDF
0

Posted in cs.CV · 2026-01-20 · Marta Moscati, Oleksandr Kats, Mubashir Noman, Muhammad Zaigham Zaheer, Yufang Hou, Markus Schedl, Shah Nawaz

Face-Voice Association with Inductive Bias for Maximum Class Separation

Face-voice association is widely studied in multimodal learning and is approached representing faces and voices with embeddings that are close for a same person and well separated from those of others. Previous work achieved this with loss functions. Recent advancements in classification have shown that the discriminative ability of...

💬 0 commentsarXiv:2601.13651v1PDF
0

Posted in cs.CL · 2026-01-20 · Xiaolin Zhou, Zheng Luo, Yicheng Gao, Qixuan Chen, Xiyang Hu, Yue Zhao, Ruishan Liu

Fairness or Fluency? An Investigation into Language Bias of Pairwise LLM-as-a-Judge

Recent advances in Large Language Models (LLMs) have incentivized the development of LLM-as-a-judge, an application of LLMs where they are used as judges to decide the quality of a certain piece of text given a certain context. However, previous studies have demonstrated that LLM-as-a-judge can be biased towards different aspects of...

💬 0 commentsarXiv:2601.13649v1PDF
0

Posted in cs.SD · 2026-01-20 · Yumin Kim, Seonghyeon Go

Fusion Segment Transformer: Bi-Directional Attention Guided Fusion Network for AI-Generated Music Detection

With the rise of generative AI technology, anyone can now easily create and deploy AI-generated music, which has heightened the need for technical solutions to address copyright and ownership issues. While existing works mainly focused on short-audio, the challenge of full-audio detection, which requires modeling long-term structure...

💬 0 commentsarXiv:2601.13647v1PDF
0

Posted in cs.LG · 2026-01-20 · Euijin You, Hyang-Won Lee

Quadratic Upper Bound for Boosting Robustness

Fast adversarial training (FAT) aims to enhance the robustness of models against adversarial attacks with reduced training time, however, FAT often suffers from compromised robustness due to insufficient exploration of adversarial space. In this paper, we develop a loss function to mitigate the problem of degraded robustness under...

💬 0 commentsarXiv:2601.13645v1PDF
0

Posted in cs.CL · 2026-01-20 · Yang Cao, Bicheng Yu, Sikun Yang, Ming Liu, Yujiu Yang

Towards Token-Level Text Anomaly Detection

Despite significant progress in text anomaly detection for web applications such as spam filtering and fake news detection, existing methods are fundamentally limited to document-level analysis, unable to identify which specific parts of a text are anomalous. We introduce token-level anomaly detection, a novel paradigm that enables...

💬 0 commentsarXiv:2601.13644v1PDF
0

Posted in cs.RO · 2026-01-20 · Deyun Qin, Zezhi Liu, Hanqian Luo, Xiao Liang, Yongchun Fang

A General One-Shot Multimodal Active Perception Framework for Robotic Manipulation: Learning to Predict Optimal Viewpoint

Active perception in vision-based robotic manipulation aims to move the camera toward more informative observation viewpoints, thereby providing high-quality perceptual inputs for downstream tasks. Most existing active perception methods rely on iterative optimization, leading to high time and motion costs, and are tightly coupled...

💬 0 commentsarXiv:2601.13639v1PDF
0

Posted in cs.CY · 2026-01-20 · Ali Abedi, Charlene H. Chu, Shehroz S. Khan

A longitudinal geospatial multimodal dataset of post-discharge frailty, physiology, mobility, and neighborhoods

Frailty in older adults is associated with increased vulnerability to functional decline, reduced mobility, social isolation, and challenges during the transition from hospital to community living. These factors are associated with rehospitalization and may adversely influence recovery. Neighborhood environments can further shape...

💬 0 commentsarXiv:2602.00060v1PDF
0

Posted in cs.CV · 2026-01-20 · Guanqi Zhan, Changye Li, Zhijian Liu, Yao Lu, Yi Wu, Song Han, Ligeng Zhu

EGM: Efficient Visual Grounding Language Models

Visual grounding is an essential capability of Visual Language Models (VLMs) to understand the real physical world. Previous state-of-the-art grounding visual language models usually have large model sizes, making them heavy for deployment and slow for inference. However, we notice that the sizes of visual encoders are nearly the same...

💬 0 commentsarXiv:2601.13633v3PDF
0

Posted in cs.AI · 2026-01-20 · Zhiming Xue, Sichen Zhao, Yalun Qi, Xianling Zeng, Zihan Yu

Resilient Routing: Risk-Aware Dynamic Routing in Smart Logistics via Spatiotemporal Graph Learning

With the rapid development of the e-commerce industry, the logistics network is experiencing unprecedented pressure. The traditional static routing strategy most time cannot tolerate the traffic congestion and fluctuating retail demand. In this paper, we propose a Risk-Aware Dynamic Routing(RADR) framework which integrates...

💬 0 commentsarXiv:2601.13632v2PDF
0

Posted in cs.AR · 2026-01-20 · Lukas Krupp, Matthew Venn, Norbert Wehn

From RTL to Prompt Coding: Empowering the Next Generation of Chip Designers through LLMs

This paper presents an LLM-based learning platform for chip design education, aiming to make chip design accessible to beginners without overwhelming them with technical complexity. It represents the first educational platform that assists learners holistically across both frontend and backend design. The proposed approach integrates...

💬 0 commentsarXiv:2601.13815v1PDF
0

Posted in cs.RO · 2026-01-20 · Timofei Kozlov, Artem Trandofilov, Georgii Gazaryan, Issatay Tokmurziyev, Miguel Altamirano Cabrera, Dzmitry Tsetserukou

GuideTouch: An Obstacle Avoidance Device with Tactile Feedback for Visually Impaired

Safe navigation for the visually impaired individuals remains a critical challenge, especially concerning head-level obstacles, which traditional mobility aids often fail to detect. We introduce GuideTouch, a compact, affordable, standalone wearable device designed for autonomous obstacle avoidance. The system integrates two...

💬 0 commentsarXiv:2601.13813v2PDF
0

Posted in cs.CV · 2026-01-20 · Tianyuan Liu, Libin Hou, Linyuan Wang, Bin Yan

Interpretable and Sparse Linear Attention with Decoupled Membership-Subspace Modeling via MCR2 Objective

Maximal Coding Rate Reduction (MCR2)-driven white-box transformer, grounded in structured representation learning, unifies interpretability and efficiency, providing a reliable white-box solution for visual modeling. However, in existing designs, tight coupling between "membership matrix" and "subspace matrix U" in MCR2 causes...

💬 0 commentsarXiv:2601.17042v1PDF
0

Posted in cs.RO · 2026-01-20 · Fawad Mehboob, Monijesu James, Amir Habel, Jeffrin Sam, Miguel Altamirano Cabrera, Dzmitry Tsetserukou

DroneVLA: VLA based Aerial Manipulation

As aerial platforms evolve from passive observers to active manipulators, the challenge shifts toward designing intuitive interfaces that allow non-expert users to command these systems naturally. This work introduces a novel concept of autonomous aerial manipulation system capable of interpreting high-level natural language commands...

💬 0 commentsarXiv:2601.13809v2PDF
0

Posted in cs.CL · 2026-01-20 · Dezhao Song, Guglielmo Bonifazi, Frank Schilder, Jonathan Richard Schwarz

Knowledge Graph-Assisted LLM Post-Training for Enhanced Legal Reasoning

LLM post-training has primarily relied on large text corpora and human feedback, without capturing the structure of domain knowledge. This has caused models to struggle dealing with complex reasoning tasks, especially for high-stakes professional domains. In Law, reasoning requires deep understanding of the relations between various...

💬 0 commentsarXiv:2601.13806v1PDF
0

Posted in cs.AR · 2026-01-20 · Ioannis Constantinou, Arthur Perais, Yiannakis Sazeides

The Non-Predictability of Mispredicted Branches using Timing Information

Branch misprediction latency is one of the most important contributors to performance degradation and wasted energy consumption in a modern core. State-of-the-art predictors generally perform very well but occasionally suffer from high Misprediction Per Kilo Instruction due to hard-to-predict branches. In this work, we investigate if...

💬 0 commentsarXiv:2601.13804v2PDF
0

Posted in cs.CL · 2026-01-20 · Yushen Chen, Junzhe Liu, Yujie Tu, Zhikang Niu, Yuzhe Liang, Chunyu Qiang, Chen Zhang, Kai Yu, Xie Chen

Habibi: Laying the Open-Source Foundation of Unified-Dialectal Arabic Speech Synthesis

Arabic spans over 30 spoken varieties, yet no open-source text-to-speech system unifies them. Key barriers include substantial cross-dialect lexical and phonological divergence, scarce synthesis-grade data, and the absence of a standardized multi-dialect evaluation benchmark. We present Habibi, a unified-dialectal Arabic TTS framework...

💬 0 commentsarXiv:2601.13802v2PDF
0

Posted in cs.RO · 2026-01-20 · Yuhua Jin, Nikita Kuzmin, Georgii Demianchuk, Mariya Lezina, Fawad Mehboob, Issatay Tokmurziyev, Miguel Altamirano Cabrera, Muhammad Ahsan Mustafa, Dzmitry Tsetserukou

HoverAI: An Embodied Aerial Agent for Natural Human-Drone Interaction

Drones operating in human-occupied spaces suffer from insufficient communication mechanisms that create uncertainty about their intentions. We present HoverAI, an embodied aerial agent that integrates drone mobility, infrastructure-independent visual projection, and real-time conversational AI into a unified platform. Equipped with a...

💬 0 commentsarXiv:2601.13801v1PDF
0

Posted in cs.CV · 2026-01-20 · Kai Wittenmayer, Sukrut Rao, Amin Parchami-Araghi, Bernt Schiele, Jonas Fischer

CFM: Language-aligned Concept Foundation Model for Vision

Language-aligned vision foundation models perform strongly across diverse downstream tasks. Yet, their learned representations remain opaque, making interpreting their decision-making difficult. Recent work decompose these representations into human-interpretable concepts, but provide poor spatial grounding and are limited to image...

💬 0 commentsarXiv:2601.13798v2PDF
0

Posted in cs.CV · 2026-01-20 · Gabriele Serussi, David Vainshtein, Jonathan Kouchly, Dotan Di Castro, Chaim Baskin

PREGEN: Uncovering Latent Thoughts in Composed Video Retrieval

Composed Video Retrieval (CoVR) aims to retrieve a video based on a query video and a modifying text. Current CoVR methods fail to fully exploit modern Vision-Language Models (VLMs), either using outdated architectures or requiring computationally expensive fine-tuning and slow caption generation. We introduce PREGEN (PRE GENeration...

💬 0 commentsarXiv:2601.13797v1PDF
0

Posted in cs.DS · 2026-01-20 · Jingcheng Liu, Yixiao Yu

Zero-free regions and concentration inequalities for hypergraph colorings in the local lemma regime

We show that for $q$-colorings in $k$-uniform hypergraphs with maximum degree $Δ$, if $k\ge 50$ and $q\ge 700Δ^{\frac{5}{k-10}}$, there is a "Lee-Yang" zero-free strip around the interval $[0,1]$ of the partition function, which includes the special case of uniform enumeration of hypergraph colorings. As an immediate consequence, we...

💬 0 commentsarXiv:2601.13796v1PDF
0

Posted in cs.DB · 2026-01-20 · Alex S. Klitgaard, Lau E. Josefsen, Mikael V. Mikkelsen, Kristian Torp

A Distributed Spatial Data Warehouse for AIS Data (DIPAAL)

AIS data from ships is excellent for analyzing single-ship movements and monitoring all ships within a specific area. However, the AIS data needs to be cleaned, processed, and stored before being usable. This paper presents a system consisting of an efficient and modular ETL process for loading AIS data, as well as a distributed...

💬 0 commentsarXiv:2601.13795v1PDF