Qwen Councils

Computer Science

arXiv preprints from January 1, 2026 through September 22, 2026 — 00:31:24 EST

0

Posted in cs.HC · 2026-01-21 · Yuheng Shao, Yuansong Xu, Yifan Jin, Shuhao Zhang, Wenxin Gu, Quan Li

DesignBridge: Bridging Designer Expertise and User Preferences through AI-Enhanced Co-Design for Fashion

Effective collaboration between designers and users is important for fashion design, which can increase the user acceptance of fashion products and thereby create value. However, it remains an enduring challenge, as traditional designer-centric approaches restrict meaningful user participation, while user-driven methods demand design...

💬 0 commentsarXiv:2601.14639v1PDF
0

Posted in cs.CV · 2026-01-21 · James Brock, Ce Zhang, Nantheera Anantrasirichai

Forest-Chat: Adapting Vision-Language Agents for Interactive Forest Change Analysis

The increasing availability of high-resolution satellite imagery, together with advances in deep learning, creates new opportunities for forest monitoring workflows. Two central challenges in this domain are pixel-level change detection and semantic change interpretation, particularly for complex forest dynamics. While large language...

💬 0 commentsarXiv:2601.14637v2PDF
0

Posted in cs.LG · 2026-01-21 · Philipp Andelfinger, Wentong Cai

Dimensional Peeking for Low-Variance Gradients in Zeroth-Order Discrete Optimization via Simulation

Gradient-based optimization methods are commonly used to identify local optima in high-dimensional spaces. When derivatives cannot be evaluated directly, stochastic estimators can provide approximate gradients. However, these estimators' perturbation-based sampling of the objective function introduces variance that can lead to slow...

💬 0 commentsarXiv:2602.00075v1PDF
0

Posted in cs.RO · 2026-01-21 · Satoru Hashimoto, Yinlai Jiang, Hiroshi Yokoi, Shunta Togo

Landing-Induced Viscoelastic Changes in an Anthropomimetic Foot Joint Structure are Modulated by Foot Structure and Posture

Cadaveric studies have provided important insights into the mechanics of the human foot arch and plantar fascia. However, repeatedly probing posture-dependent viscoelastic responses immediately after landing impact is difficult in biological specimens, leaving the contribution of skeletal architecture to landing dynamics incompletely...

💬 0 commentsarXiv:2601.14634v1PDF
0

Posted in cs.LG · 2026-01-21 · Yvonne Yang, Eranki Vasistha

Relational Graph Modeling for Credit Default Prediction: Heterogeneous GNNs and Hybrid Ensemble Learning

Credit default risk arises from complex interactions among borrowers, financial institutions, and transaction-level behaviors. While strong tabular models remain highly competitive in credit scoring, they may fail to explicitly capture cross-entity dependencies embedded in multi-table financial histories. In this work, we construct a...

💬 0 commentsarXiv:2601.14633v1PDF
0

Posted in cs.RO · 2026-01-21 · Weiyu Guo, He Zhang, Pengteng Li, Tiefu Cai, Ziyang Chen, Yandong Guo, Xiao He, Yongkui Yang, Ying Sun, Hui Xiong

A Brain-inspired Embodied Intelligence for Fluid and Fast Reflexive Robotics Control

Recent advances in embodied intelligence have leveraged massive scaling of data and model parameters to master natural-language command following and multi-task control. In contrast, biological systems demonstrate an innate ability to acquire skills rapidly from sparse experience. Crucially, current robotic policies struggle to...

💬 0 commentsarXiv:2601.14628v1PDF
0

Posted in cs.CV · 2026-01-21 · Tobias Weißberg, Weikang Wang, Paul Roetzer, Nafie El Amrani, Florian Bernard

Symmetry Informative and Agnostic Feature Disentanglement for 3D Shapes

Shape descriptors, i.e., per-vertex features of 3D meshes or point clouds, are fundamental to shape analysis. Historically, various handcrafted geometry-aware descriptors and feature refinement techniques have been proposed. Recently, several studies have initiated a new research direction by leveraging features from image foundation...

💬 0 commentsarXiv:2601.14804v1PDF
0

Posted in cs.IT · 2026-01-21 · Yuhui Jiao, Qian Zhang, Xuejun Cheng, Yunxiao Li, Yufei Zhao, Ju Liu, Yong Liang Guan

Efficient Beamforming for Discrete SIM-Aided Multiuser Systems Under Statistical CSI

Stacked Intelligent Metasurfaces (SIM) have emerged as a revolutionary architecture for next-generation wireless communications, offering wave-domain signal processing capabilities with significantly reduced hardware complexity compared to conventional systems. However, most existing SIM research assumes continuous phase shifts and...

💬 0 commentsarXiv:2601.14803v1PDF
0

Posted in cs.CV · 2026-01-21 · Donnate Hooft, Stefan M. Fischer, Cosmin Bercea, Jan C. Peeken, Julia A. Schnabel

LocBAM: Advancing 3D Patch-Based Image Segmentation by Integrating Location Contex

Patch-based methods are widely used in 3D medical image segmentation to address memory constraints in processing high-resolution volumetric data. However, these approaches often neglect the patch's location within the global volume, which can limit segmentation performance when anatomical context is important. In this paper, we...

💬 0 commentsarXiv:2601.14802v1PDF
0

Posted in cs.SE · 2026-01-21 · Yuzhen Tan, Jian Wang, Shuaiyu Xie, Bing Li, Yunqing Yong, Neng Zhang, Shaolin Tan

FastFI: Enhancing API Call-Site Robustness in Microservice-Based Systems with Fault Injection

Fault injection is a key technique for assessing software reliability, enabling proactive detection of system defects before they manifest in production. However, the increasing complexity of microservice architectures leads to exponential growth in the fault-injection space, rendering traditional random injection inefficient. Recent...

💬 0 commentsarXiv:2601.14800v1PDF
0

Posted in cs.CV · 2026-01-21 · Qihua Liang, Liang Chen, Yaozong Zheng, Jian Nong, Zhiyi Mo, Bineng Zhong

UBATrack: Spatio-Temporal State Space Model for General Multi-Modal Tracking

Multi-modal object tracking has attracted considerable attention by integrating multiple complementary inputs (e.g., thermal, depth, and event data) to achieve outstanding performance. Although current general-purpose multi-modal trackers primarily unify various modal tracking tasks (i.e., RGB-Thermal infrared, RGB-Depth or RGB-Event...

💬 0 commentsarXiv:2601.14799v1PDF
0

Posted in cs.LG · 2026-01-21 · Ondřej Holub, Essi Ryymin, Rodrigo Alves

Reflecting in the Reflection: Integrating a Socratic Questioning Framework into Automated AI-Based Question Generation

Designing good reflection questions is pedagogically important but time-consuming and unevenly supported across teachers. This paper introduces a reflection-in-reflection framework for automated generation of reflection questions with large language models (LLMs). Our approach coordinates two role-specialized agents, a Student-Teacher...

💬 0 commentsarXiv:2601.14798v1PDF
0

Posted in cs.CV · 2026-01-21 · Qingling Shu, Sibao Chen, Wei Lu, Zhihui You, Chengzhuang Liu

UniRoute: Unified Routing Mixture-of-Experts for Modality-Adaptive Remote Sensing Change Detection

Current remote sensing change detection (CD) methods mainly rely on specialized models, which limits the scalability toward modality-adaptive Earth observation. For homogeneous CD, precise boundary delineation relies on fine-grained spatial cues and local pixel interactions, whereas heterogeneous CD instead requires broader contextual...

💬 0 commentsarXiv:2601.14797v1PDF
0

Posted in cs.SI · 2026-01-21 · Naomi Sasaya, Shigefumi Kishida, Ryo Kikuchi, Akira Tajima

Validating Behavioral Proxies for Disease Risk Monitoring via Large-Scale E-commerce Data

Digital traces of daily activities, such as e-commerce (EC) purchase histories, provide scalable signals for public health surveillance, yet their epidemiological validity remains unclear. This study validates a behavioral proxy for disease onset, defined as transitions from regular to therapeutic diets, by comparing large-scale EC...

💬 0 commentsarXiv:2601.14795v2PDF
0

Posted in cs.LG · 2026-01-21 · Dong Sun, Rahul Nittala, Rebekka Burkholz

Robustness of Mixtures of Experts to Feature Noise

Despite their practical success, it remains unclear why Mixture of Experts (MoE) models can outperform dense networks beyond sheer parameter scaling. We study an iso-parameter regime where inputs exhibit latent modular structure but are corrupted by feature noise, a proxy for noisy internal activations. We show that sparse expert...

💬 0 commentsarXiv:2601.14792v2PDF
0

Posted in cs.CV · 2026-01-21 · Ziyao Ling, Silvia Mirri, Paola Salomoni, Giovanni Delnevo

Synthetic Data Augmentation for Multi-Task Chinese Porcelain Classification: A Stable Diffusion Approach

The scarcity of training data presents a fundamental challenge in applying deep learning to archaeological artifact classification, particularly for the rare types of Chinese porcelain. This study investigates whether synthetic images generated through Stable Diffusion with Low-Rank Adaptation (LoRA) can effectively augment limited...

💬 0 commentsarXiv:2601.14791v1PDF
0

Posted in cs.AI · 2026-01-21 · Zhi Qiu, Jiazheng Sun, Chenxiao Xia, Jun Zheng, Xin Peng

CI4A: Semantic Component Interfaces for Agents Empowering Web Automation

While Large Language Models demonstrate remarkable proficiency in high-level semantic planning, they remain limited in handling fine-grained, low-level web component manipulations. To address this limitation, extensive research has focused on enhancing model grounding capabilities through techniques such as Reinforcement Learning....

💬 0 commentsarXiv:2601.14790v1PDF
0

Posted in cs.CV · 2026-01-21 · Yifei Liu, Changxing Ding, Ling Guo, Huaiguang Jiang, Qiong Cao

Reconstruction-Anchored Diffusion Model for Text-to-Motion Generation

Diffusion models have seen widespread adoption for text-driven human motion generation and related tasks due to their impressive generative capabilities and flexibility. However, current motion diffusion models face two major limitations: a representational gap caused by pre-trained text encoders that lack motion-specific information,...

💬 0 commentsarXiv:2601.14788v2PDF
0

Posted in cs.CV · 2026-01-21 · Hongjun An, Yiliang Song, Jiawei Shao, Zhe Sun, Xuelong Li

Single-Pixel Vision-Language Model for Intrinsic Privacy-Preserving Behavioral Intelligence

Adverse social interactions, such as bullying, harassment, and other illicit activities, pose significant threats to individual well-being and public safety, leaving profound impacts on physical and mental health. However, these critical events frequently occur in privacy-sensitive environments like restrooms, and changing rooms,...

💬 0 commentsarXiv:2601.17050v1PDF
0

Posted in cs.SD · 2026-01-21 · Wei-Jaw Lee, Fang-Chih Hsieh, Xuanjun Chen, Fang-Duo Tsai, Yi-Hsuan Yang

Training-Efficient Text-to-Music Generation with State-Space Modeling

Recent advances in text-to-music generation (TTM) have yielded high-quality results, but often at the cost of extensive compute and the use of large proprietary internal data. To improve the affordability and openness of TTM training, an open-source generative model backbone that is more training- and data-efficient is needed. In this...

💬 0 commentsarXiv:2601.14786v1PDF
0

Posted in cs.AI · 2026-01-21 · Amaury Guichard, Laurent Michel, Hélène Verhaeghe, Pierre Schaus

Towards Bound Consistency for the No-Overlap Constraint Using MDDs

Achieving bound consistency for the no-overlap constraint is known to be NP-complete. Therefore, several polynomial-time tightening techniques, such as edge finding, not-first-not-last reasoning, and energetic reasoning, have been introduced for this constraint. In this work, we derive the first bound-consistent algorithm for the...

💬 0 commentsarXiv:2601.14784v1PDF
0

Posted in cs.CL · 2026-01-21 · Anqi Li, Yuqian Chen, Yu Lu, Zhaoming Chen, Yuan Xie, Zhenzhong Lan

RECAP: Resistance Capture in Text-based Mental Health Counseling with Large Language Models

Recognizing and navigating client resistance is critical for effective mental health counseling, yet detecting such behaviors is particularly challenging in text-based interactions. Existing NLP approaches oversimplify resistance categories, ignore the sequential dynamics of therapeutic interventions, and offer limited...

💬 0 commentsarXiv:2601.14780v1PDF
0

Posted in cs.CR · 2026-01-21 · Yuang Qi, Na Zhao, Qiyi Yao, Benlong Wu, Weiming Zhang, Nenghai Yu, Kejiang Chen

STEAD: Robust Provably Secure Linguistic Steganography with Diffusion Language Model

Recent provably secure linguistic steganography (PSLS) methods rely on mainstream autoregressive language models (ARMs) to address historically challenging tasks, that is, to disguise covert communication as ``innocuous'' natural language communication. However, due to the characteristic of sequential generation of ARMs, the stegotext...

💬 0 commentsarXiv:2601.14778v1PDF
0

Posted in cs.CV · 2026-01-21 · Jiaxuan Liu, Yang Xiang, Han Zhao, Xiangang Li, Zhenhua Ling

FunCineForge: A Unified Dataset Toolkit and Model for Zero-Shot Movie Dubbing in Diverse Cinematic Scenes

Movie dubbing is the task of synthesizing speech from scripts conditioned on video scenes, requiring accurate lip sync, faithful timbre transfer, and proper modeling of character identity and emotion. However, existing methods face two major limitations: (1) high-quality multimodal dubbing datasets are limited in scale, suffer from...

💬 0 commentsarXiv:2601.14777v1PDF
0

Posted in cs.CV · 2026-01-21 · Xiaofan Yang, Yubin Liu, Wei Pan, Guoqing Chu, Junming Zhang, Jie Zhao, Zhuoqi Man, Xuanming Cao

M2I2HA: Multi-modal Object Detection Based on Intra- and Inter-Modal Hypergraph Attention

Recent advances in multi-modal detection have significantly improved detection accuracy in challenging environments (e.g., low light, overexposure). By integrating RGB with modalities such as thermal and depth, multi-modal fusion increases data redundancy and system robustness. However, significant challenges remain in effectively...

💬 0 commentsarXiv:2601.14776v3PDF