Qwen Councils

All arXiv

arXiv preprints from January 1, 2026 through July 21, 2026 — 19:16:39 EST

0

Posted in eess.AS · 2026-01-20 · Aditya Kamlesh Parikh, Cristian Tejedor-Garcia, Catia Cucchiarini, Helmer Strik

Zero-Shot Speech LLMs for Multi-Aspect Evaluation of L2 Speech: Challenges and Opportunities

An accurate assessment of L2 English pronunciation is crucial for language learning, as it provides personalized feedback and ensures a fair evaluation of individual progress. However, automated scoring remains challenging due to the complexity of sentence-level fluency, prosody, and completeness. This paper evaluates the zero-shot...

💬 0 commentsarXiv:2601.16230v1PDF
0

Posted in cs.IT · 2026-01-20 · Maria Abu-Sini, Reinhard Heckel

Near Optimal Code Construction for the Adversarial Torn Paper Channel With Edit Errors

Motivated by DNA storage systems and 3D fingerprinting, this work studies the adversarial torn paper channel with edit errors. This channel first applies at most $t_e$ edit errors (i.e., insertions, deletions, and substitutions) to the transmitted word and then breaks it into $t+1$ fragments at arbitrary positions. In this paper, we...

💬 0 commentsarXiv:2601.14088v1PDF
0

Posted in cs.AR · 2026-01-20 · Ruichi Han, Yizhi Chen, Tong Lei, Jordi Altayo Gonzalez, Ahmed Hemani

'1'-bit Count-based Sorting Unit to Reduce Link Power in DNN Accelerators

Interconnect power consumption remains a bottleneck in Deep Neural Network (DNN) accelerators. While ordering data based on '1'-bit counts can mitigate this via reduced switching activity, practical hardware sorting implementations remain underexplored. This work proposes the hardware implementation of a comparison-free sorting unit...

💬 0 commentsarXiv:2601.14087v1PDF
0

Posted in cs.CV · 2026-01-20 · Nattapong Kurpukdee, Adrian G. Bors

Two-Stream temporal transformer for video action classification

Motion representation plays an important role in video understanding and has many applications including action recognition, robot and autonomous guidance or others. Lately, transformer networks, through their self-attention mechanism capabilities, have proved their efficiency in many applications. In this study, we introduce a new...

💬 0 commentsarXiv:2601.14086v1PDF
0

Posted in math.AP · 2026-01-20 · Seyed Abdolhamid Banihashemi, Huy Q. Nguyen

On large periodic traveling wave solutions to the free boundary Stokes and Navier-Stokes equations

We study the free boundary problem for a finite-depth layer of viscous incompressible fluid in arbitrary dimension, modeled by the Stokes or Navier-Stokes equations. In addition to the gravitational field acting in the bulk, the free boundary is acted upon by surface tension and an external stress tensor posited to be in traveling...

💬 0 commentsarXiv:2601.14085v1PDF
0

Posted in cs.CV · 2026-01-20 · Abdurrahim Yilmaz, Ozan Erdem, Ece Gokyayla, Ayda Acar, Burc Bugra Dagtas, Dilara Ilhan Erdil, Gulsum Gencoglan, Burak Temelkuran

DermaBench: A Clinician-Annotated Benchmark Dataset for Dermatology Visual Question Answering and Reasoning

Vision-language models (VLMs) are increasingly important in medical applications; however, their evaluation in dermatology remains limited by datasets that focus primarily on image-level classification tasks such as lesion recognition. While valuable for recognition, such datasets cannot assess the full visual understanding, language...

💬 0 commentsarXiv:2601.14084v1PDF
0

Posted in astro-ph.CO · 2026-01-20 · Sofia Contarini, Giovanni Verza, Alice Pisani

The era of precision cosmology with voids

Cosmic voids, the large underdense regions of our Universe, have emerged over the past decade as powerful cosmological laboratories: their simple dynamics, sensitivity to local gravitational effects and cosmic expansion, and ability to span large volumes, make them uniquely suited to test fundamental physics. Fueled by advances in...

💬 0 commentsarXiv:2601.14362v1PDF
0

Posted in hep-ph · 2026-01-20 · Claudia Cornella, Max Ferré, Matthias König, Matthias Neubert

The Simplest B Decay, Precisely

We derive the QCD$\times$QED factorization theorem governing the leptonic decay $B^-\toμ^-\barν_μ(γ)$ at all orders in $α_s$ and $α$. Electromagnetic corrections to this decay probe multiple scales, which we disentangle through a sequence of effective field theories (EFTs). The resulting state-of-the-art prediction for the...

💬 0 commentsarXiv:2601.14361v2PDF
0

Posted in astro-ph.HE · 2026-01-20 · O. S. Salafia, A. Celotti, E. Sobacchi, L. Nava, G. Oganesyan, G. Ghirlanda, S. Boula, M. E. Ravasio, G. Ghisellini

A self-consistent explanation of the MeV line in GRB 221009A unveils a dense circum-stellar medium

GRB~221009A has been the brightest gamma-ray burst (GRB) observed to date, and its afterglow has been characterized with unprecedented detail at TeV energies by LHAASO. Quite puzzlingly, it is also the most energetic GRB known. Among the riddles posed by this mysterious source, however, the sheer energetics are hardly the most...

💬 0 commentsarXiv:2601.14257v2PDF
0

Posted in cs.CV · 2026-01-20 · Matthew Gwilliam, Xiao Wang, Xuefeng Hu, Zhenheng Yang

Implicit Neural Representation Facilitates Unified Universal Vision Encoding

Models for image representation learning are typically designed for either recognition or generation. Various forms of contrastive learning help models learn to convert images to embeddings that are useful for classification, detection, and segmentation. On the other hand, models can be trained to reconstruct images with pixel-wise,...

💬 0 commentsarXiv:2601.14256v1PDF
0

Posted in cs.CV · 2026-01-20 · Sangbeom Lim, Seoung Wug Oh, Jiahui Huang, Heeji Yoon, Seungryong Kim, Joon-Young Lee

VideoMaMa: Mask-Guided Video Matting via Generative Prior

Generalizing video matting models to real-world videos remains a significant challenge due to the scarcity of labeled data. To address this, we present Video Mask-to-Matte Model (VideoMaMa) that converts coarse segmentation masks into pixel accurate alpha mattes, by leveraging pretrained video diffusion models. VideoMaMa demonstrates...

💬 0 commentsarXiv:2601.14255v1PDF
0

Posted in astro-ph.EP · 2026-01-20 · James G. Rogers, James E. Owen, Ethan Schreyer, James Kirk

Using observations of escaping H/He to constrain the atmospheric composition of sub-Neptunes

The internal composition of sub-Neptunes remains a prominent unresolved question in exoplanetary science. We present a technique to place constraints on envelope mean molecular weight that utilises observations of escaping hydrogen or helium exospheres. This method is based on a simple timescale argument, which states that...

💬 0 commentsarXiv:2601.14254v2PDF
0

Posted in cs.CV · 2026-01-20 · Hongyuan Chen, Xingyu Chen, Youjia Zhang, Zexiang Xu, Anpei Chen

Motion 3-to-4: 3D Motion Reconstruction for 4D Synthesis

We present Motion 3-to-4, a feed-forward framework for synthesising high-quality 4D dynamic objects from a single monocular video and an optional 3D reference mesh. While recent advances have significantly improved 2D, video, and 3D content generation, 4D synthesis remains difficult due to limited training data and the inherent...

💬 0 commentsarXiv:2601.14253v1PDF
0

Posted in q-bio.NC · 2026-01-20 · Cole Korponay

A Dual-Head Transformer-State-Space Architecture for Neurocircuit Mechanism Decomposition from fMRI

Precision psychiatry aspires to elucidate brain-based biomarkers of psychopathology to bolster disease risk assessment and treatment development. To this end, functional magnetic resonance imaging (fMRI) has helped triangulate brain circuits whose functional features are correlated with or even predictive of forms of psychopathology....

💬 0 commentsarXiv:2601.15344v1PDF
0

Posted in cs.IT · 2026-01-20 · Tristan Simas

Semantic Identity Compression: Zero-Error Laws, Rate-Distortion, and Neurosymbolic Necessity

Symbolic systems operate over precise identities: variables denote specific objects, pointers target precise memory locations, and database keys refer to singular records. Neural embeddings generalize by compressing away semantic detail, but this compression creates collision ambiguity: multiple distinct entities can share the same...

💬 0 commentsarXiv:2601.14252v6PDF
0

Posted in cs.CV · 2026-01-20 · Said Taghadouini, Adrien Cavaillès, Baptiste Aubertin

LightOnOCR: A 1B End-to-End Multilingual Vision-Language Model for State-of-the-Art OCR

We present LightOnOCR-2-1B, a 1B-parameter end-to-end multilingual vision--language model that converts document images (e.g., PDFs) into clean, naturally ordered text without brittle OCR pipelines. Trained on a large-scale, high-quality distillation mix with strong coverage of scans, French documents, and scientific PDFs,...

💬 0 commentsarXiv:2601.14251v2PDF
0

Posted in cs.CV · 2026-01-20 · Pengze Zhang, Yanze Wu, Mengtian Li, Xu Bai, Songtao Zhao, Fulong Ye, Chong Mou, Xinghui Li, Zhuowei Chen, Qian He, Mingyuan Gao

OmniTransfer: All-in-one Framework for Spatio-temporal Video Transfer

Videos convey richer information than images or text, capturing both spatial and temporal dynamics. However, most existing video customization methods rely on reference images or task-specific temporal priors, failing to fully exploit the rich spatio-temporal information inherent in videos, thereby limiting flexibility and...

💬 0 commentsarXiv:2601.14250v1PDF
0

Posted in cs.CL · 2026-01-20 · Yuming Yang, Mingyoung Lai, Wanxu Zhao, Xiaoran Fan, Zhiheng Xi, Mingqi Wu, Chiyue Huang, Jun Zhao, Haijun Lv, Jian Tong, Yunhua Zhou, Yicheng Zou, Qipeng Guo, Tao Gui, Qi Zhang, Xuanjing Huang

Which Reasoning Trajectories Teach Students to Reason Better? A Simple Metric of Informative Alignment

Long chain-of-thought (CoT) trajectories provide rich supervision signals for distilling reasoning from teacher to student LLMs. However, both prior work and our experiments show that trajectories from stronger teachers do not necessarily yield better students, highlighting the importance of data-student suitability in distillation....

💬 0 commentsarXiv:2601.14249v5PDF
0

Posted in astro-ph.GA · 2026-01-20 · Tehreem N. Hai, Kristen B. W. McQuinn, Yao-Yuan Mao, Roger E. Cohen, David Shih, Erik Tollerud, Joseph A. Breneman, Andrew E. Dolphin, Max J. B. Newman, Adam Smercina

A Quenched and Relatively Isolated Dwarf Galaxy in the Local Volume

An increasing number of discoveries of isolated and quenched dwarf galaxies are challenging the idea that the present-day local environment of low-mass systems is the main determinant of their quenching. We present new Hubble Space Telescope (HST) data of one such system, the dwarf galaxy Canes Venatici C (CVn C). CVn C is a low-mass...

💬 0 commentsarXiv:2601.14248v1PDF
0

Posted in math.DS · 2026-01-20 · Murilo R. Cândido, Douglas D. Novaes, Joan S. G. Rivera

Detecting Limit Tori in Non-Smooth Systems: An Analytic Approach with Applications to 3D Piecewise Linear Systems

This work investigates a class of non-autonomous $T$-periodic piecewise smooth differential systems and their associated time-$T$ maps. Our main result provides an analytical approach for detecting, within this class of piecewise differential systems, isolated invariant tori associated with normally hyperbolic invariant closed curves...

💬 0 commentsarXiv:2601.14247v1PDF
0

Posted in cs.CV · 2026-01-20 · Zeyuan Chen, Kai Zhang, Zhuowen Tu, Yuanjun Xiong

Soft Tail-dropping for Adaptive Visual Tokenization

We present Soft Tail-dropping Adaptive Tokenizer (STAT), a 1D discrete visual tokenizer that adaptively chooses the number of output tokens per image according to its structural complexity and level of detail. STAT encodes an image into a sequence of discrete codes together with per-token keep probabilities. Beyond standard...

💬 0 commentsarXiv:2601.14246v1PDF
0

Posted in cs.IR · 2026-01-20 · Zhongyu Yang, Wei Pang, Yingfang Yuan

XR: Cross-Modal Agents for Composed Image Retrieval

Retrieval is being redefined by agentic AI, demanding multimodal reasoning beyond conventional similarity-based paradigms. Composed Image Retrieval (CIR) exemplifies this shift as each query combines a reference image with textual modifications, requiring compositional understanding across modalities. While embedding-based CIR methods...

💬 0 commentsarXiv:2601.14245v2PDF
0

Posted in astro-ph.IM · 2026-01-20 · Alice P. Curtin, Reshma Anna-Thomas, Amanda M. Cook, Carolina Cruz-Vinaccia, Jason Hessels, Robert Main, Inés Pastor Marazuela, Lauren Rhodes, Vishwangi Shah

One Attempt at Building an Inclusive & Accessible Hybrid Astronomy Conference: FRB 2025

The rapid expansion of the Fast Radio Burst (FRB) field has been accompanied by a simultaneous growth of FRB conferences. While these meetings are essential for interacting with other researchers and establishing collaborations, many remain only accessible to those with substantial travel funding, flexible schedules, or geographical...

💬 0 commentsarXiv:2601.14357v2PDF
0

Posted in eess.SP · 2026-01-20 · Qing Zhang, Adham Sakhnini, Robbert Beerten, Haoqiu Xiong, Zhuangzhuang Cui, Yang Miao, Sofie Pollin

Robust Localization in OFDM-Based Massive MIMO through Phase Offset Calibration

Accurate localization in Orthogonal Frequency Division Multiplexing (OFDM)-based massive Multiple-Input Multiple-Output (MIMO) systems depends critically on phase coherence across subcarriers and antennas. However, practical systems suffer from frequency-dependent and (spatial) antenna-dependent phase offsets, degrading localization...

💬 0 commentsarXiv:2601.14244v1PDF
0

Posted in cs.LG · 2026-01-20 · Haocheng Xi, Charlie Ruan, Peiyuan Liao, Yujun Lin, Han Cai, Yilong Zhao, Shuo Yang, Kurt Keutzer, Song Han, Ligeng Zhu

Jet-RL: Enabling On-Policy FP8 Reinforcement Learning with Unified Training and Rollout Precision Flow

Reinforcement learning (RL) is essential for enhancing the complex reasoning capabilities of large language models (LLMs). However, existing RL training pipelines are computationally inefficient and resource-intensive, with the rollout phase accounting for over 70% of total training time. Quantized RL training, particularly using FP8...

💬 0 commentsarXiv:2601.14243v2PDF