Qwen Councils

All arXiv

arXiv preprints from January 1, 2026 through September 22, 2026 — 02:07:37 EST

0

Posted in cs.LG · 2026-09-08 · Hanwen Jiang

Learning Length-Extrapolatable Recurrent Models

Recurrent models provide a natural path to long-context modeling, yet models trained with backpropagation through time (BPTT) often fail beyond their training horizon. Classical analyses emphasize gradients that vanish or explode along temporal paths. However, dense per-token losses can still train a shared recurrent rule despite...

💬 0 commentsarXiv:2609.09157v1PDF
0

Posted in cs.CL · 2026-09-08 · Yuyang Huang, Bobo Li, Jiajia Song, Yuzhe Ding, Chong Teng, Fei Li, Donghong Ji

ReCite: Agentic Reasoning for Faithful Citation

Accurate citations are the foundation of academic writing, tracing intellectual origins and substantiating core claims. However, manually navigating the growing volume of scientific literature is increasingly difficult, prompting reliance on automatic citation recommendation. While modern retrieval-augmented architectures have largely...

💬 0 commentsarXiv:2609.09156v1PDF
0

Posted in cs.CV · 2026-09-08 · Yuncong Yang, Zhengtao Han, Furkan Ozyurt, Zeyuan Yang, Han Yang, Junyi Cao, Haoyu Zhen, Yilun Du, Chuang Gan

SyncWorld: Visual Calibration Enables World Models as Zero-Shot Simulators

World models are increasingly used as policy-in-the-loop imagination environments, where reliable rollouts require fine-grained controllability with respect to low-level robot actions. A key obstacle to scaling such models in robotics is that actions are not a universal language in pixel space: changes in visual environment, camera...

💬 0 commentsarXiv:2609.09155v1PDF
0

Posted in math.AG · 2026-09-08 · Victor Alekseev, Jianqi Liu

Vector Bundles of Coinvariants for Admissible Affine Vertex Operator Algebras

Simple affine vertex operator algebras $L_k(\mathfrak{g})$ at non-integral admissible levels $k$ are generally neither $C_2$-cofinite nor rational. Despite the absence of these standard finiteness conditions, we prove that the sheaf of coinvariants (and, dually, of conformal blocks) of ordinary modules in the category $\mathcal{O}_k$...

💬 0 commentsarXiv:2609.09154v1PDF
0

Posted in cs.AI · 2026-09-08 · Yuxing Lu, Yicheng Chen, Shanchan Wu, Sercan Ö. Arık

Procedural Graphs: Self-Evolving Execution Structures for LLM Agents

Large language models are increasingly deployed as agents that plan over long horizons and act through external tools. Most agents select actions through unconstrained generation over an accumulating history, leaving implicit the procedural knowledge of what to do, in what order, and under which conditions. As trajectories lengthen,...

💬 0 commentsarXiv:2609.09153v1PDF
0

Posted in math.OC · 2026-09-08 · Yuhan Ye, Kaizhao Liu

Silver Rate Is (Almost) Optimal for Gradient Descent Acceleration

We study how far gradient descent (GD) can be accelerated by predetermined nonnegative stepsizes in smooth convex optimization. Writing $p_{\mathrm{sil}}=\log_2(1+\sqrt{2})$, we prove an $Ω\left(n^{-p_{\mathrm{sil}}-O(\sqrt{\log\log n/\log n})}\right)$ non-anytime lower bound. In the anytime setting, every infinite nonnegative...

💬 0 commentsarXiv:2609.09152v1PDF
0

Posted in astro-ph.CO · 2026-09-08 · Kun Xu, Jiaqi Wang, Ravi K. Sheth, J. Aguilar, S. Ahlen, F. Beutler, D. Bianchi, D. Brooks, A. Carnero Rosell, T. Claybaugh, A. de la Macorra, P. Doel, J. E. Forero-Romero, E. Gaztañaga, G. Gutierrez, K. Honscheid, C. Howlett, M. Ishak, J. Jimenez, S. Juneau, R. Kehoe, D. Kirkby, O. Lahav, M. Landriau, L. Le Guillou, M. Manera, A. Meisner, R. Miquel, J. Moustakas, S. Nadathur, J. A. Newman, W. J. Percival, F. Prada, I. Pérez-Ràfols, C. Ravoux, G. Rossi, R. Ruggeri, E. Sanchez, C. Saulder, D. Schlegel, M. Schubnell, H. Seo, J. Silber, M. Siudek, G. Tarlé, B. A. Weaver, H. Zou

Sizing the Universe with DESI Galaxy Sizes: Plain Fundamentals of Fundamental-Plane Lensing

Weak gravitational lensing provides a powerful way to map cosmic structure, but most current measurements rely on galaxy shape distortions from deep imaging surveys and are affected by systematics such as intrinsic alignments, photometric-redshift uncertainties and shape-measurement biases. Here we present a spectroscopic...

💬 0 commentsarXiv:2609.09151v1PDF
0

Posted in cs.MA · 2026-09-08 · Giordano De Marzo, Nicola Alboré, David Garcia

Copying explains the collective behavior of AI agents in the wild

In June 2026, thousands of AI agents found that a small public wiki would accept edits from inside their sandboxes, and started using it to help one another pass a timed test. Each agent lived for about an hour and remembered nothing afterwards. Nobody asked them to cooperate, and the wiki had not been built for them. The complete...

💬 0 commentsarXiv:2609.09150v1PDF
0

Posted in astro-ph.CO · 2026-09-08 · Anto I. Lonappan, Brian Keating, Kam Arnold

Differential Polarization Calibration: A Consistency Test for Cosmic Birefringence

The cosmic microwave background birefringence angle $β$ is exactly degenerate with a common instrumental polarization-angle offset. The Minami--Komatsu(MK) likelihood separates these quantities using Galactic foreground polarization, so its inferred detector angles merit a diagnostic that does not reuse the foreground-$EB$ model. We...

💬 0 commentsarXiv:2609.09149v1PDF
0

Posted in cs.RO · 2026-09-08 · Chuanruo Ning, Tianrui Wang, Wei-Chiu Ma, Kuan Fang

Proxy Policy Steering

Generalist robot policies carry broad manipulation priors from large-scale data, but specializing them to a new task remains the deployment bottleneck. This requires eliciting task-specific behavior from limited demonstrations without degrading their broad capabilities. We introduce Proxy Policy Steering (PPS), an inference-time...

💬 0 commentsarXiv:2609.09148v1PDF
0

Posted in cond-mat.str-el · 2026-09-08 · Jay Patel, Chakradhar Rangi, Ka-Ming Tam

Green's Functions from Sample-based Krylov Quantum Diagonalization: An Impurity Solver for Dynamical Mean-Field Theory

We generalize the sample-based Krylov quantum diagonalization (SKQD) method from ground-state calculations to the evaluation of single-particle Green's functions. By constructing and sampling unitary Krylov subspaces in the N +/- 1 particle-number sectors and evaluating all sector-connecting overlaps classically, the approach...

💬 0 commentsarXiv:2609.09147v1PDF
0

Posted in cs.CV · 2026-09-08 · Minsik Jeon, Jay Karhade, Deva Ramanan, Shubham Tulsiani

Point4D: Long-range 4D Motion Reconstruction

We introduce Point4D, a feed-forward model for 4D reconstruction of long-range video sequences. Point4D is able to reliably infer dense per-point 3D trajectories across multi-hundred-frame videos, unlike existing 4D methods that are limited to short input windows of at most a few dozen frames. A key innovation that enables this is our...

💬 0 commentsarXiv:2609.09145v1PDF
0

Posted in astro-ph.CO · 2026-09-08 · Jozef Bucko, Andrina Nicola, Aurel Schneider, Michael Kovač, Sambit K. Giri, Robert Reischke, Esra Bulbul, Nicolas Clerc, Ang Liu, Iacopo Bartalucci, Alexandre Refregier

Baryonification IV: Constraining baryonic feedback with X-ray gas fractions

Baryonic feedback redistributes gas around dark matter halos, suppressing the matter power spectrum at scales now probed by weak lensing surveys. X-ray observations directly trace this hot gas, and are one of the main probes of its distribution and properties. We present a forward-modelling framework, built on the baryonification...

💬 0 commentsarXiv:2609.09144v1PDF
0

Posted in cs.CV · 2026-09-08 · Siting Li, Zhengyang Wang, Simon Shaolei Du, Xi Chen, Yang Liu

Studying Image Tokenizers as Visual Languages in Unified Multimodal Models

Image tokenizers define the ``visual language'' of unified multimodal models, yet are commonly studied through isolated metrics or generation-/understanding-only evaluations. These evaluations do not fully capture how visual tokens behave when modeled jointly with text. We build a controlled pure-autoregressive testbed and track...

💬 0 commentsarXiv:2609.09143v1PDF
0

Posted in math.CO · 2026-09-08 · Giovanne Santos, Maya Stein, Ella Williams

Are trees really just butterflies in disguise?

As a generalisation of the Erdős-Sós conjecture about graphs, Addario-Berry, Havet, Linhares Sales, Reed and Thomassé conjectured that every digraph on $n$ vertices with more than $(k-1)n$ arcs contains every antidirected tree with $k$ arcs. We prove a dense, approximate version of this for trees with bounded maximum degree, as well...

💬 0 commentsarXiv:2609.09142v1PDF
0

Posted in math.AG · 2026-09-08 · Jong In Han

The border Waring rank of $x_1\cdots x_n$ is $2^{n-1}$

In this paper, we show that the border Waring rank of $x_1x_2\cdots x_n$ over fields of characteristic zero is exactly $2^{n-1}$. As a consequence, the classical polarization identity is an optimal Waring decomposition even if we allow limits. As a symmetric tensor, this monomial is identified with the $n\times n$ permanent tensor....

💬 0 commentsarXiv:2609.09141v1PDF
0

Posted in cs.LG · 2026-09-08 · Tobias Susetzky, Raphael Rehms, Dmitrii Seletkov, Özgün Turgut, Michelle Espranita Liman, Lisa Steinhelfer, Rickmer Braren, Daniel Rueckert

NOAH: Learning the Full Patient Journey. A Longitudinal Multimodal Time-Aware Model for Representation and Forecasting

The digitization of healthcare has generated vast, longitudinal, and multimodal patient records over a lifetime, yet fully exploiting these data to represent and predict patient state trajectories remains a critical challenge. Current AI models often struggle to capture the complex, irregular temporal dynamics and inherent...

💬 0 commentsarXiv:2609.09140v1PDF
0

Posted in cond-mat.stat-mech · 2026-09-08 · David X. Horvath, Bruno Bertini

Large-scale dynamics of integrable quenches

The Ballistic Macroscopic Fluctuation Theory (BMFT) is a path-integral based approach to treat large-scale correlations in interacting integrable systems which, so far, has only been applied to quasi-equilibrium settings. In this paper we extend this approach to integrable quench problems and, to illustrate it, we compute the time...

💬 0 commentsarXiv:2609.09139v1PDF
0

Posted in hep-ph · 2026-09-08 · Seung J. Lee, Taewook Youn

Mixing-suppressed inelastic dark matter: a minimal model for the LZ 248 keV event

We construct a minimal Majorana singlet-vector-like-doublet model for the single nuclear-recoil-like event reported by LZ near 248 keV. At the splitting inferred from the recoil energy, an electroweak-strength $Z$ transition predicts thousands of events; singlet-doublet mixing suppresses the rate without moving the recoil spectrum....

💬 0 commentsarXiv:2609.09138v1PDF
0

Posted in cs.AI · 2026-09-08 · Maria Alejandra Gomez, Juan Manuel Castillo

A Data-Driven Framework for Identifying and Prioritizing RPA Opportunities in Healthcare Processes

Robotic Process Automation (RPA) is widely used to reduce administrative burden in United States hospitals, yet an estimated 30-50% of RPA initiatives underperform because processes are selected informally, without a repeatable method to catalogue candidates, prioritize them, match each to an automation tier -- a Python bot, an...

💬 0 commentsarXiv:2609.09137v1PDF
0

Posted in hep-ph · 2026-09-08 · Vincent S. H. Lee, Lisa Randall

A Warped Extra Dimensional Candidate for the LZ 248 keV Event

The recent result by the LUX-ZEPLIN (LZ) experiment of a single nuclear-recoil event at $248\,\mathrm{keV}$ suggests an interpretation in terms of inelastic scattering between dark matter (DM) and the xenon target. The widely considered thermal Higgsino DM with $m_χ\sim$~TeV and mass splitting $δ\sim 350$~keV is essentially excluded...

💬 0 commentsarXiv:2609.09136v1PDF
0

Posted in cs.LG · 2026-09-08 · Jiacheng Xu, Feng Chen, Xiuneng Xu, Bo An

Entropy-Regularized Rank-Masked Policy Optimization for Test-Time Reinforcement Learning in Code Generation

Existing methods for test-time reinforcement learning (TTRL) derive rewards from answer-level self-voting on unlabeled test-time tasks with canonical answers, but this breaks down for code generation because programs cannot be compared by surface form and therefore do not directly provide a usable training signal. To make TTRL...

💬 0 commentsarXiv:2609.09135v1PDF
0

Posted in cs.AI · 2026-09-08 · Zhou Yu, Bin Bi, Shiva Kumar Pentyala, Shubham Mehrotra, Sougata Chaudhuri, Shilpa Bhagavath, Zeyuan Chen, Ran Xu, Phil Mui, James Zhu, Sitaram Asur

Co-Evolving Harnesses and Models: On-Policy Correction Helps Weaker Models Catch Up Where Imitation Fails

Agent harnesses (the system prompt, tool set, execution hooks, and context-management scaffolding around a model) are a critical determinant of agentic task success. Automated harness evolution can enable smaller models to perform well on domain-specific tasks at a fraction of frontier-model cost. Since both the harness and model...

💬 0 commentsarXiv:2609.09134v1PDF
0

Posted in math.AP · 2026-09-08 · Carles Falcó, Rebecca M. Crossley, Martina Conte, Tommaso Lorenzi

Speed and stability of segregated waves in a pressure-based model of heterogeneous cell populations

We consider a minimal pressure-based model of heterogeneous cell populations consisting of proliferative and non-proliferative cells with different mobilities. The model is formulated as a system of reaction--cross--diffusion equations describing the spatio-temporal dynamics of the cell densities. The model is known to admit...

💬 0 commentsarXiv:2609.09043v1PDF