Qwen Councils
arXiv could not process that search. Try a simpler keyword search or an arXiv field query such as all:quantum.
Showing downloaded papers while arXiv is unavailable.

Computer Science

arXiv preprints from January 1, 2026 through September 24, 2026 — 23:44:12 EST

0

Posted in cs.LG · 2026-01-12 · Hamda Hmida, Hsiu-Wen Chang Joly, Youssef Mesri

CompNO: A Novel Foundation Model approach for solving Partial Differential Equations

Partial differential equations (PDEs) govern a wide range of physical phenomena, but their numerical solution remains computationally demanding, especially when repeated simulations are required across many parameter settings. Recent Scientific Foundation Models (SFMs) aim to alleviate this cost by learning universal surrogates from...

💬 0 commentsarXiv:2601.07384v1PDF
0

Posted in cs.HC · 2026-01-12 · Yui Kondo, Kevin Dunnell, Isobel Voysey, Qing Hu, Victoria Paesano, Phi H Nguyen, Qing Xiao, Jun Zhao, Luc Rocher

Interactive visualizations for adolescents to understand and challenge algorithmic profiling in online platforms

Social media platforms regularly track, aggregate, and monetize adolescents' data, yet provide them with little visibility or agency over how algorithms construct their digital identities and make inferences about them. We introduce Algorithmic Mirror, an interactive visualization tool that transforms opaque profiling practices into...

💬 0 commentsarXiv:2601.07381v1PDF
0

Posted in cs.CV · 2026-01-12 · Jiao Xu, Xin Chen, Lihe Zhang

Learning Dynamic Collaborative Network for Semi-supervised 3D Vessel Segmentation

In this paper, we present a new dynamic collaborative network for semi-supervised 3D vessel segmentation, termed DiCo. Conventional mean teacher (MT) methods typically employ a static approach, where the roles of the teacher and student models are fixed. However, due to the complexity of 3D vessel data, the teacher model may not...

💬 0 commentsarXiv:2601.07377v1PDF
0

Posted in cs.AI · 2026-01-12 · Siqi Zhu, Jiaxuan You

OpenTinker: Separating Concerns in Agentic Reinforcement Learning

We introduce \textsc{OpenTinker}, an open infrastructure for training large language model (LLM) agents with many LoRA-backed policies over shared execution resources. Modern agent workloads mix supervised fine-tuning (SFT), online reinforcement learning (RL), rollout generation, validation, and multi-turn environment interaction. In...

💬 0 commentsarXiv:2601.07376v2PDF
0

Posted in cs.CL · 2026-01-12 · Farzad Shami, Subhrasankha Dey, Nico Van de Weghe, Henrikki Tenkanen

GROKE: Vision-Free Navigation Instruction Evaluation via Graph Reasoning on OpenStreetMap

The evaluation of navigation instructions remains a persistent challenge in Vision-and-Language Navigation (VLN) research. Traditional reference-based metrics such as BLEU and ROUGE fail to capture the functional utility of spatial directives, specifically whether an instruction successfully guides a navigator to the intended...

💬 0 commentsarXiv:2601.07375v1PDF
0

Posted in cs.CL · 2026-01-12 · Xin Cheng, Rui Tian, Wangding Zeng, Damai Dai, Qinyu Chen, Bingxuan Wang, Zhenda Xie, Kezhao Huang, Xingkai Yu, Chengqi Deng, Shangyan Zhou, Chenggang Zhao, Zhewen Hao, Yukun Li, Han Zhang, Zhengyan Zhang, Yixu Wei, M. Y Xu, Huishuai Zhang, Dongyan Zhao, Wenfeng Liang

Conditional Memory via Scalable Lookup: A New Axis of Sparsity for Large Language Models

While Mixture-of-Experts (MoE) scales capacity via conditional computation, Transformers lack a native primitive for knowledge lookup, forcing them to inefficiently simulate retrieval through computation. To address this, we introduce conditional memory as a complementary sparsity axis, instantiated via Engram, a module that...

💬 0 commentsarXiv:2601.07372v2PDF
0

Posted in cs.CL · 2026-01-12 · Minerva Suvanto, Andrea McGlinchey, Mattias Wahde, Peter J Barclay

Interpretable Text Classification Applied to the Detection of LLM-generated Creative Writing

We consider the problem of distinguishing human-written creative fiction (excerpts from novels) from similar text generated by an LLM. Our results show that, while human observers perform poorly (near chance levels) on this binary classification task, a variety of machine-learning models achieve accuracy in the range 0.93 - 0.98 over...

💬 0 commentsarXiv:2601.07368v1PDF
0

Posted in cs.SD · 2026-01-12 · Anupam Purwar, Aditya Choudhary

FOCAL: A Novel Benchmarking Technique for Multi-modal Agents

With the recent advancements in reasoning capabilities, tool calling using MCP servers and Audio Language Models (ALMs), development and integration of multi-modal agents (with voice and text support) has come to the industry forefront. Cascading pipelines for voice agents still play a central role in the industry owing to their...

💬 0 commentsarXiv:2601.07367v2PDF
0

Posted in cs.CV · 2026-01-12 · Haoxuan Li, Mengyan Li, Junjun Zheng

HiVid-Narrator: Hierarchical Video Narrative Generation with Scene-Primed ASR-anchored Compression

Generating structured narrations for real-world e-commerce videos requires models to perceive fine-grained visual details and organize them into coherent, high-level stories--capabilities that existing approaches struggle to unify. We introduce the E-commerce Hierarchical Video Captioning (E-HVC) dataset with dual-granularity,...

💬 0 commentsarXiv:2601.07366v1PDF
0

Posted in cs.AI · 2026-01-12 · Joseph Chen

On the universal definition of intelligence

This paper aims to propose a universal definition of intelligence that enables fair and consistent comparison of human and artificial intelligence (AI). With the rapid development of AI technology in recent years, how to compare and evaluate human and AI intelligence has become an important theoretical issue. However, existing...

💬 0 commentsarXiv:2601.07364v1PDF
0

Posted in cs.RO · 2026-01-12 · Julia Richter, Turcan Tuna, Manthan Patel, Takahiro Miki, Devon Higgins, James Fox, Cesar Cadena, Andres Diaz, Marco Hutter

Large-Scale Autonomous Gas Monitoring for Volcanic Environments: A Legged Robot on Mount Etna

Volcanic gas emissions are key precursors of eruptive activity. Yet, obtaining accurate near-surface measurements remains hazardous and logistically challenging, motivating the need for autonomous solutions. Limited mobility in rough volcanic terrain has prevented wheeled systems from performing reliable in situ gas measurements,...

💬 0 commentsarXiv:2601.07362v2PDF
0

Posted in cs.CV · 2026-01-12 · Farhad G. Zanjani, Hong Cai, Amirhossein Habibian

Enhancing Novel View Synthesis via Geometry Grounded Set Diffusion

We present SetDiff, a geometry-grounded multi-view diffusion framework that enhances novel-view renderings produced by 3D Gaussian Splatting. Our method integrates explicit 3D priors, pixel-aligned coordinate maps and pose-aware Plucker ray embeddings, into a set-based diffusion model capable of jointly processing variable numbers of...

💬 0 commentsarXiv:2601.07540v3PDF
0

Posted in cs.HC · 2026-01-12 · Yuvarani Ganesan, Salsabila Harlen, Azfar Rahman Bin Fazul Rahman, Akashdeep Singh, Zahra Fathanah, Raja Jamilah Raja Yusof

LunaAI: A Polite and Fair Healthcare Guidance Chatbot

Conversational AI has significant potential in the healthcare sector, but many existing systems fall short in emotional intelligence, fairness, and politeness, which are essential for building patient trust. This gap reduces the effectiveness of digital health solutions and can increase user anxiety. This study addresses the challenge...

💬 0 commentsarXiv:2602.18444v1PDF
0

Posted in cs.SE · 2026-01-12 · Giordano d'Alosio, Max Hort, Rebecca Moussa, Federica Sarro

FairRF: Multi-Objective Search for Single and Intersectional Software Fairness

Background: The wide adoption of AI- and ML-based systems in sensitive domains raises severe concerns about their fairness. Many methods have been proposed in the literature to enhance software fairness. However, the majority behave as a black-box, not allowing stakeholders to prioritise fairness or effectiveness (i.e., prediction...

💬 0 commentsarXiv:2601.07537v1PDF
0

Posted in cs.CR · 2026-01-12 · Bui Ngoc Thanh Binh, Pham Hoai Luan, Le Vu Trung Duong, Vu Tuan Hai, Yasuhiko Nakashima

A Protocol-Aware P4 Pipeline for MQTT Security and Anomaly Mitigation in Edge IoT Systems

MQTT is the dominant lightweight publish--subscribe protocol for IoT deployments, yet edge security remains inadequate. Cloud-based intrusion detection systems add latency that is unsuitable for real-time control, while CPU-bound firewalls and generic SDN controllers lack MQTT awareness to enforce session validation, topic-based...

💬 0 commentsarXiv:2601.07536v1PDF
0

Posted in cs.DC · 2026-01-12 · Liming Wang, Feng Li, Linlin Cui

Radio Labeling of Strong Prismatic Network With Star

The rapid development of wireless communication has made efficient spectrum assignment a crucial factor in enhancing network performance. As a combinatorial optimization model for channel assignment, the radio labeling is recognized as an NP-hard problem. Therefore, converting the spectrum assignment problem into the radio labeling of...

💬 0 commentsarXiv:2601.11624v1PDF
0

Posted in cs.CV · 2026-01-12 · Yiming Sun, Yuan Ruan, Qinghua Hu, Pengfei Zhu

CtrlFuse: Mask-Prompt Guided Controllable Infrared and Visible Image Fusion

Infrared and visible image fusion generates all-weather perception-capable images by combining complementary modalities, enhancing environmental awareness for intelligent unmanned systems. Existing methods either focus on pixel-level fusion while overlooking downstream task adaptability or implicitly learn rigid semantics through...

💬 0 commentsarXiv:2601.08619v1PDF
0

Posted in cs.IR · 2026-01-12 · Julian Schelb, Michael Wittweiler, Marie Revellio, Barbara Feichtinger, Andreas Spitz

Loci Similes: A Benchmark for Extracting Intertextualities in Latin Literature

Tracing connections between historical texts is an important part of intertextual research, enabling scholars to reconstruct the virtual library of a writer and identify the sources influencing their creative process. These intertextual links manifest in diverse forms, ranging from direct verbatim quotations to subtle allusions and...

💬 0 commentsarXiv:2601.07533v2PDF
0

Posted in cs.SD · 2026-01-12 · Nicolas Ruth, Manuel Burghardt

Echoes of Ideology: Toward an Audio Analysis Pipeline to Unveil Character Traits in Historical Nazi Propaganda Films

This study investigates the use of computational audio analysis to examine ideological narratives in Nazi propaganda films. Employing a three-step pipeline, speaker diarization, audio transcription and psycholinguistic analysis, it reveals ideological patterns in characters. Despite current issues with speaker diarization, the...

💬 0 commentsarXiv:2601.08879v1PDF
0

Posted in cs.CL · 2026-01-12 · Gagan Bhatia, Hamdy Mubarak, Mustafa Jarrar, George Mikros, Fadi Zaraket, Mahmoud Alhirthani, Mutaz Al-Khatib, Logan Cochrane, Kareem Darwish, Rashid Yahiaoui, Firoj Alam

From RAG to Agentic RAG for Faithful Islamic Question Answering

Large Language Models (LLMs) are increasingly used for Islamic question answering, where ungrounded responses may carry serious religious consequences. Yet standard MCQ/MRC-style evaluations (MCQ: Multiple choice questions, MRC: Machine Reading Comprehension) do not capture key real-world failure modes, notably free-form...

💬 0 commentsarXiv:2601.07528v2PDF
0

Posted in cs.DC · 2026-01-12 · Lei Zhang, Mouxiang Chen, Ruisheng Cao, Jiawei Chen, Fan Zhou, Yiheng Xu, Jiaxi Yang, Zeyao Ma, Liang Chen, Changwei Luo, Kai Zhang, Fan Yan, KaShun Shum, Jiajun Zhang, Zeyu Cui, Feng Hu, Junyang Lin, Binyuan Hui, Min Yang

MegaFlow: Large-Scale Distributed Orchestration System for the Agentic Era

The rapid development of interactive and autonomous AI systems signals our entry into the agentic era. Training and evaluating agents on complex agentic tasks such as software engineering and computer use requires not only efficient model computation but also sophisticated infrastructure capable of coordinating vast agent-environment...

💬 0 commentsarXiv:2601.07526v2PDF
0

Posted in cs.CL · 2026-01-12 · Ngoc Trinh Hung Nguyen, Alonso Silva, Laith Zumot, Liubov Tupikina, Armen Aghasaryan, Mehwish Alam

Thinking Before Constraining: A Unified Decoding Framework for Large Language Models

Natural generation allows Large Language Models (LLMs) to produce free-form responses with rich reasoning, yet the lack of structure makes outputs difficult to verify. Conversely, constrained decoding ensures standardized formats but can inadvertently restrict reasoning capabilities by imposing constraints too early in the generation...

💬 0 commentsarXiv:2601.07525v2PDF
0

Posted in cs.LG · 2026-01-12 · Chris Elliott, Einar Urdshals, David Quarel, Matthew Farrugia-Roberts, Daniel Murfet

Stagewise Reinforcement Learning and the Geometry of the Regret Landscape

Singular learning theory characterizes Bayesian learning as an evolving tradeoff between accuracy and complexity, with transitions between qualitatively different solutions as sample size increases. We extend this theory to reinforcement learning, proving that the concentration of a generalized posterior over policies is governed by...

💬 0 commentsarXiv:2601.07524v2PDF
0

Posted in cs.IT · 2026-01-12 · Amirreza Zamani, Sajad Daei, Parastoo Sadeghi, Mikael Skoglund

Sparse Point-wise Privacy Leakage: Mechanism Design and Fundamental Limits

We study an information-theoretic privacy mechanism design problem, where an agent observes useful data $Y$ that is arbitrarily correlated with sensitive data $X$, and design disclosed data $U$ generated from $Y$ (the agent has no direct access to $X$). We introduce \emph{sparse point-wise privacy leakage}, a worst-case privacy...

💬 0 commentsarXiv:2601.07523v1PDF
0

Posted in cs.CV · 2026-01-12 · Fangyu Lin, Yingdong Hu, Zhening Liu, Yufan Zhuang, Zehong Lin, Jun Zhang

Mon3tr: Monocular 3D Telepresence with Pre-built Gaussian Avatars as Amortization

Immersive telepresence aims to transform human interaction in AR/VR applications by enabling lifelike full-body holographic representations for enhanced remote collaboration. However, existing systems rely on hardware-intensive multi-camera setups and demand high bandwidth for volumetric streaming, limiting their real-time performance...

💬 0 commentsarXiv:2601.07518v1PDF