Qwen Councils
arXiv could not process that search. Try a simpler keyword search or an arXiv field query such as all:quantum.
Showing downloaded papers while arXiv is unavailable.

Computer Science

arXiv preprints from January 1, 2026 through September 24, 2026 — 20:35:31 EST

0

Posted in cs.CL · 2026-01-13 · Daniil Gurgurov, Yusser Al Ghussin, Tanja Baeumel, Cheng-Ting Chou, Patrick Schramowski, Marius Mosbach, Josef van Genabith, Simon Ostermann

CLaS-Bench: A Cross-Lingual Alignment and Steering Benchmark

Understanding and controlling the behavior of large language models (LLMs) is an increasingly important topic in multilingual NLP. Beyond prompting or fine-tuning, , i.e.,~manipulating internal representations during inference, has emerged as a more efficient and interpretable technique for adapting models to a target language. Yet,...

💬 0 commentsarXiv:2601.08331v1PDF
0

Posted in cs.CR · 2026-01-13 · Mingqi Lv, Shanshan Zhang, Haiwen Liu, Tieming Chen, Tiantian Zhu

APT-MCL: An Adaptive APT Detection System Based on Multi-View Collaborative Provenance Graph Learning

Advanced persistent threats (APTs) are stealthy and multi-stage, making single-point defenses (e.g., malware- or traffic-based detectors) ill-suited to capture long-range and cross-entity attack semantics. Provenance-graph analysis has become a prominent approach for APT detection. However, its practical deployment is hampered by (i)...

💬 0 commentsarXiv:2601.08328v1PDF
0

Posted in cs.RO · 2026-01-13 · Gabriele Calzolari, Vidya Sumathy, Christoforos Kanellakis, George Nikolakopoulos

Safe Heterogeneous Multi-Agent RL with Communication Regularization for Coordinated Target Acquisition

This paper introduces a decentralized multi-agent reinforcement learning framework enabling structurally heterogeneous teams of agents to jointly discover and acquire randomly located targets in environments characterized by partial observability, communication constraints, and dynamic interactions. Each agent's policy is trained with...

💬 0 commentsarXiv:2601.08327v1PDF
0

Posted in cs.IT · 2026-01-13 · Emil Björnson, Amna Irshad, Özlem Tugfe Demir, Giuseppe Thadeu Freitas de Abreu, Alva Kosasih, Vitaly Petrov

From Antenna Abundance to Antenna Intelligence in 6G Gigantic MIMO Systems

Current cellular systems achieve high spectral efficiency through Massive MIMO, which leverages an abundance of antennas to create favorable propagation conditions for multiuser spatial multiplexing. Looking towards future networks, the extrapolation of this paradigm leads to systems with many hundreds of antennas per base station,...

💬 0 commentsarXiv:2601.08326v1PDF
0

Posted in cs.RO · 2026-01-13 · Zhenyang Liu, Yongchong Gu, Yikai Wang, Xiangyang Xue, Yanwei Fu

ActiveVLA: Injecting Active Perception into Vision-Language-Action Models for Precise 3D Robotic Manipulation

Recent advances in robot manipulation have leveraged pre-trained vision-language models (VLMs) and explored integrating 3D spatial signals into these models for effective action prediction, giving rise to the promising vision-language-action (VLA) paradigm. However, most existing approaches overlook the importance of active...

💬 0 commentsarXiv:2601.08325v1PDF
0

Posted in cs.AI · 2026-01-13 · Yupeng Huo, Yaxi Lu, Zhong Zhang, Haotian Chen, Yankai Lin

AtomMem : Learnable Dynamic Agentic Memory with Atomic Memory Operation

Equipping agents with memory is essential for solving real-world long-horizon problems. However, most existing agent memory mechanisms rely on static and hand-crafted workflows. This limits the performance and generalization ability of these memory designs, which highlights the need for a more flexible, learning-based memory...

💬 0 commentsarXiv:2601.08323v3PDF
0

Posted in cs.NI · 2026-01-13 · Aymen Hasan Alawadi

Streamlined Pathway (SP) Approach: An Efficient Load Balancer to Enhance Quality of Service

Efficient load-balancing mechanisms are critical for maximizing performance and increasing the quality of service (QoS) of data center networks (DCNs). Obtaining the optimal QoS while minimizing resource consumption remains a significant challenge. This paper proposes the streamlined pathway (SP) model, which is a flow scheduling...

💬 0 commentsarXiv:2601.08887v1PDF
0

Posted in cs.CV · 2026-01-13 · Lichen Ma, Xiaolong Fu, Gaojing Zhou, Zipeng Guo, Ting Zhu, Yichun Liu, Yu Shi, Jason Li, Junshi Huang

UM-Text: A Unified Multimodal Model for Image Understanding and Visual Text Editing

With the rapid advancement of image generation, visual text editing using natural language instructions has received increasing attention. The main challenge of this task is to fully understand the instruction and reference image, and thus generate visual text that is style-consistent with the image. Previous methods often involve...

💬 0 commentsarXiv:2601.08321v3PDF
0

Posted in cs.CV · 2026-01-13 · Dapinder Kaur, Neeraj Battish, Arnav Bhavsar, Shashi Poddar

YOLOBirDrone: Dataset for Bird vs Drone Detection and Classification and a YOLO based enhanced learning architecture

The use of aerial drones for commercial and defense applications has benefited in many ways and is therefore utilized in several different application domains. However, they are also increasingly used for targeted attacks, posing a significant safety challenge and necessitating the development of drone detection systems. Vision-based...

💬 0 commentsarXiv:2601.08319v1PDF
0

Posted in cs.LG · 2026-01-13 · Tomoki Kubo, Ryuken Uda, Yusuke Iida

Deep Exploration of Epoch-wise Double Descent in Noisy Data: Signal Separation, Large Activation, and Benign Overfitting

Deep double descent is one of the key phenomena underlying the generalization capability of deep learning models. In this study, epoch-wise double descent, which is delayed generalization following overfitting, was empirically investigated by focusing on the evolution of internal structures. Fully connected neural networks of three...

💬 0 commentsarXiv:2601.08316v1PDF
0

Posted in cs.CV · 2026-01-13 · Kang Fu, Huiyu Duan, Zicheng Zhang, Yucheng Zhu, Jun Zhao, Xiongkuo Min, Jia Wang, Guangtao Zhai

Enhancing Image Quality Assessment Ability of LMMs via Retrieval-Augmented Generation

Large Multimodal Models (LMMs) have recently shown remarkable promise in low-level visual perception tasks, particularly in Image Quality Assessment (IQA), demonstrating strong zero-shot capability. However, achieving state-of-the-art performance often requires computationally expensive fine-tuning methods, which aim to align the...

💬 0 commentsarXiv:2601.08311v1PDF
0

Posted in cs.LG · 2026-01-13 · Kun Liang, Clive Bai, Xin Xu, Chenming Tang, Sanwoo Lee, Weijie Liu, Saiyong Yang, Yunfang Wu

ORBIT: On-policy Exploration-Exploitation for Controllable Multi-Budget Reasoning

Recent Large Reasoning Models (LRMs) achieve strong performance by leveraging long-form Chain-of-Thought (CoT) reasoning, but uniformly applying overlong reasoning at inference time incurs substantial and often unnecessary computational cost. To address this, prior work explores various strategies to infer an appropriate reasoning...

💬 0 commentsarXiv:2601.08310v2PDF
0

Posted in cs.CL · 2026-01-13 · Bo Yang, Yu Zhang, Yunkui Chen, Lanfei Feng, Xiao Xu, Nueraili Aierken, Shijian Li

AgriAgent: Contract-Driven Planning and Capability-Aware Tool Orchestration in Real-World Agriculture

Intelligent agent systems in real-world agricultural scenarios must handle diverse tasks under multimodal inputs, ranging from lightweight information understanding to complex multi-step execution. However, most existing approaches rely on a unified execution paradigm, which struggles to accommodate large variations in task complexity...

💬 0 commentsarXiv:2601.08308v1PDF
0

Posted in cs.AI · 2026-01-13 · Ethan Zhang

Uncovering Latent Bias in LLM-Based Emergency Department Triage Through Proxy Variables

Recent advances in large language models (LLMs) have enabled their integration into clinical decision-making; however, hidden biases against patients across racial, social, economic, and clinical backgrounds persist. In this study, we investigate bias in LLM-based medical AI systems applied to emergency department (ED) triage. We...

💬 0 commentsarXiv:2601.15306v1PDF
0

Posted in cs.CV · 2026-01-13 · Dongting Hu, Aarush Gupta, Magzhan Gabidolla, Arpit Sahni, Huseyin Coskun, Yanyu Li, Yerlan Idelbayev, Ahsan Mahmood, Aleksei Lebedev, Dishani Lahiri, Anujraaj Goyal, Ju Hu, Mingming Gong, Sergey Tulyakov, Anil Kag

SnapGen++: Unleashing Diffusion Transformers for Efficient High-Fidelity Image Generation on Edge Devices

Recent advances in diffusion transformers (DiTs) have set new standards in image generation, yet remain impractical for on-device deployment due to their high computational and memory costs. In this work, we present an efficient DiT framework tailored for mobile and edge devices that achieves transformer-level generation quality under...

💬 0 commentsarXiv:2601.08303v3PDF
0

Posted in cs.CL · 2026-01-13 · Marvin Schmitt, Anne Schwerk, Sebastian Lempert

Enhancing Sentiment Classification and Irony Detection in Large Language Models through Advanced Prompt Engineering Techniques

This study investigates the use of prompt engineering to enhance large language models (LLMs), specifically GPT-4o-mini and gemini-1.5-flash, in sentiment analysis tasks. It evaluates advanced prompting techniques like few-shot learning, chain-of-thought prompting, and self-consistency against a baseline. Key tasks include sentiment...

💬 0 commentsarXiv:2601.08302v1PDF
0

Posted in cs.CV · 2026-01-13 · Qizhen Lan, Yu-Chun Hsu, Nida Saddaf Khan, Xiaoqian Jiang

ReCo-KD: Region- and Context-Aware Knowledge Distillation for Efficient 3D Medical Image Segmentation

Accurate 3D medical image segmentation is vital for diagnosis and treatment planning, but state-of-the-art models are often too large for clinics with limited computing resources. Lightweight architectures typically suffer significant performance loss. To address these deployment and speed constraints, we propose Region- and...

💬 0 commentsarXiv:2601.08301v1PDF
0

Posted in cs.CL · 2026-01-13 · Tony Cristofano

Surgical Refusal Ablation: Disentangling Safety from Intelligence via Concept-Guided Spectral Cleaning

Safety-aligned language models systematically refuse harmful requests. While activation steering can modulate refusal, ablating the raw "refusal vector" calculated from contrastive harmful and harmless prompts often causes collateral damage and distribution drift. We argue this degradation occurs because the raw vector is...

💬 0 commentsarXiv:2601.08489v1PDF
0

Posted in cs.RO · 2026-01-13 · Chong Zhang, Victor Klemm, Fan Yang, Marco Hutter

AME-2: Agile and Generalized Legged Locomotion via Attention-Based Neural Map Encoding

Achieving agile and generalized legged locomotion across terrains requires tight integration of perception and control, especially under occlusions and sparse footholds. Existing methods have demonstrated agility on parkour courses but often rely on end-to-end sensorimotor models with limited generalization and interpretability. By...

💬 0 commentsarXiv:2601.08485v2PDF
0

Posted in cs.CV · 2026-01-13 · MD Fatin Ishraque Ayon, Sabrin Nahar, Ataur Rahman, Md. Taslim Arif, Abdul Hasib, A. S. M. Ahsanul Sarkar Akib

An IoT-Enabled Smart Aquarium System for Real-Time Water Quality Monitoring and Automated Feeding

Maintaining optimal water quality in aquariums is critical for aquatic health but remains challenging due to the need for continuous monitoring of multiple parameters. Traditional manual methods are inefficient, labor-intensive, and prone to human error, often leading to suboptimal aquatic conditions. This paper presents an IoT-based...

💬 0 commentsarXiv:2601.08484v1PDF
0

Posted in cs.LG · 2026-01-13 · Chenxu Han, Sean Bin Yang, Jilin Hu

DiffMM: Efficient Method for Accurate Noisy and Sparse Trajectory Map Matching via One Step Diffusion

Map matching for sparse trajectories is a fundamental problem for many trajectory-based applications, e.g., traffic scheduling and traffic flow analysis. Existing methods for map matching are generally based on Hidden Markov Model (HMM) or encoder-decoder framework. However, these methods continue to face significant challenges when...

💬 0 commentsarXiv:2601.08482v1PDF
0

Posted in cs.CR · 2026-01-13 · Aryan Pasikhani, Prosanta Gope, Yang Yang, Shagufta Mehnaz, Biplab Sikdar

Baiting AI: Deceptive Adversary Against AI-Protected Industrial Infrastructures

This paper explores a new cyber-attack vector targeting Industrial Control Systems (ICS), particularly focusing on water treatment facilities. Developing a new multi-agent Deep Reinforcement Learning (DRL) approach, adversaries craft stealthy, strategically timed, wear-out attacks designed to subtly degrade product quality and reduce...

💬 0 commentsarXiv:2601.08481v1PDF
0

Posted in cs.CL · 2026-01-13 · Francesco Dettori, Matteo Forasassi, Lorenzo Veronese, Livia Lestingi, Vincenzo Scotti, Matteo Giovanni Rossi

Do You Understand How I Feel?: Towards Verified Empathy in Therapy Chatbots

Conversational agents are increasingly used as support tools along mental therapeutic pathways with significant societal impacts. In particular, empathy is a key non-functional requirement in therapeutic contexts, yet current chatbot development practices provide no systematic means to specify or verify it. This paper envisions a...

💬 0 commentsarXiv:2601.08477v1PDF
0

Posted in cs.CV · 2026-01-13 · Hao Tang, Yu Liu, Shuanglin Yan, Fei Shen, Shengfeng He, Jing Qin

Cross-modal Proxy Evolving for OOD Detection with Vision-Language Models

Reliable zero-shot detection of out-of-distribution (OOD) inputs is critical for deploying vision-language models in open-world settings. However, the lack of labeled negatives in zero-shot OOD detection necessitates proxy signals that remain effective under distribution shift. Existing negative-label methods rely on a fixed set of...

💬 0 commentsarXiv:2601.08476v2PDF
0

Posted in cs.AI · 2026-01-13 · JungMin Yun, Juhwan Choi, Kyohoon Jin, Soojin Jang, Jinhee Jang, YoungBin Kim

SUMMPILOT: Bridging Efficiency and Customization for Interactive Summarization System

This paper incorporates the efficiency of automatic summarization and addresses the challenge of generating personalized summaries tailored to individual users' interests and requirements. To tackle this challenge, we introduce SummPilot, an interaction-based customizable summarization system. SummPilot leverages a large language...

💬 0 commentsarXiv:2601.08475v1PDF