Qwen Councils

Computer Science

arXiv preprints from January 1, 2026 through July 28, 2026 — 19:12:07 EST

0

Posted in cs.CV · 2026-01-03 · Hyeonjeong Ha, Jinjin Ge, Bo Feng, Kaixin Ma, Gargi Chakraborty

NarrativeTrack: Evaluating Entity-Centric Reasoning for Narrative Understanding

Multimodal large language models (MLLMs) have achieved impressive progress in vision-language reasoning, yet their ability to understand temporally unfolding narratives in videos remains underexplored. True narrative understanding requires grounding who is doing what, when, and where, maintaining coherent entity representations across...

💬 0 commentsarXiv:2601.01095v3PDF
0

Posted in cs.HC · 2026-01-03 · Yubo Shu, Peng Zhang, Meng Wu, Yan Chen, Haoxuan Zhou, Guanming Liu, Yu Zhang, Liuxin Zhang, Qianying Wang, Tun Lu, Ning Gu

SoulSeek: Exploring the Use of Social Cues in LLM-based Information Seeking

Social cues, which convey others' presence, behaviors, or identities, play a crucial role in human information seeking by helping individuals judge relevance and trustworthiness. However, existing LLM-based search systems primarily rely on semantic features, creating a misalignment with the socialized cognition underlying natural...

💬 0 commentsarXiv:2601.01094v1PDF
0

Posted in cs.CL · 2026-01-03 · Haq Nawaz Malik

ks-lit-3m: A 3.1 million word kashmiri text dataset for large language model pretraining

Large Language Models (LLMs) demonstrate remarkable fluency across high-resource languages yet consistently fail to generate coherent text in Kashmiri, a language spoken by approximately seven million people. This performance disparity stems not from inherent model limitations but from a critical scarcity of high-quality training...

💬 0 commentsarXiv:2601.01091v1PDF
0

Posted in cs.CV · 2026-01-03 · NAVER Cloud HyperCLOVA X Team

HyperCLOVA X 32B Think

In this report, we present HyperCLOVA X 32B Think, a vision-language model designed with particular emphasis on reasoning within the Korean linguistic and cultural context, as well as agentic ability. HyperCLOVA X 32B Think is pre-trained with a strong focus on reasoning capabilities and subsequently post-trained to support multimodal...

💬 0 commentsarXiv:2601.03286v1PDF
0

Posted in cs.CV · 2026-01-03 · Wangyuan Zhu, Jun Yu

Multimodal Sentiment Analysis based on Multi-channel and Symmetric Mutual Promotion Feature Fusion

Multimodal sentiment analysis is a key technology in the fields of human-computer interaction and affective computing. Accurately recognizing human emotional states is crucial for facilitating smooth communication between humans and machines. Despite some progress in multimodal sentiment analysis research, numerous challenges remain....

💬 0 commentsarXiv:2601.02415v1PDF
0

Posted in cs.MA · 2026-01-03 · Erica Coppolillo, Luca Luceri, Emilio Ferrara

Harm in AI-Driven Societies: An Audit of Toxicity Adoption on Chirper.ai

Large Language Models (LLMs) are increasingly embedded in autonomous agents that engage, converse, and co-evolve in online social platforms. While prior work has documented the generation of toxic content by LLMs, far less is known about how exposure to harmful content shapes agent behavior over time, particularly in environments...

💬 0 commentsarXiv:2601.01090v2PDF
0

Posted in cs.LG · 2026-01-03 · Nobuyuki Ota

Central Dogma Transformer: Towards Mechanism-Oriented AI for Cellular Understanding

Understanding cellular mechanisms requires integrating information across DNA, RNA, and protein - the three molecular systems linked by the Central Dogma of molecular biology. While domain-specific foundation models have achieved success for each modality individually, they remain isolated, limiting our ability to model integrated...

💬 0 commentsarXiv:2601.01089v2PDF
0

Posted in cs.CV · 2026-01-03 · Haq Nawaz Malik

600k-ks-ocr: a large-scale synthetic dataset for optical character recognition in kashmiri script

This technical report presents the 600K-KS-OCR Dataset, a large-scale synthetic corpus comprising approximately 602,000 word-level segmented images designed for training and evaluating optical character recognition systems targeting Kashmiri script. The dataset addresses a critical resource gap for Kashmiri, an endangered Dardic...

💬 0 commentsarXiv:2601.01088v1PDF
0

Posted in cs.CV · 2026-01-03 · Jichao Zhu, Jun Yu

MIAR: Modality Interaction and Alignment Representation Fuison for Multimodal Emotion

Multimodal Emotion Recognition (MER) aims to perceive human emotions through three modes: language, vision, and audio. Previous methods primarily focused on modal fusion without adequately addressing significant distributional differences among modalities or considering their varying contributions to the task. They also lacked robust...

💬 0 commentsarXiv:2601.02414v1PDF
0

Posted in cs.NI · 2026-01-03 · Jianpeng Qi, Chao Liu, Chengrui Wang, Rui Wang, Junyu Dong, Yanwei Yu

Decision-Aware Semantic State Synchronization in Compute-First Networking

In Compute-First Networking (CFN), an Access Point (AP) makes task offloading decisions based on resource state information reported by a Service Node (SN). A fundamental challenge arises from the trade-off between update overhead and decision accuracy: Frequent state updates consume limited network resources, while infrequent updates...

💬 0 commentsarXiv:2601.01086v1PDF
0

Posted in cs.CV · 2026-01-03 · Jiayi Xu, Zhang Zhang, Yuanrui Zhang, Ruitao Chen, Yixian Xu, Tianyu He, Di He

Luminark: Training-free, Probabilistically-Certified Watermarking for General Vision Generative Models

In this paper, we introduce \emph{Luminark}, a training-free and probabilistically-certified watermarking method for general vision generative models. Our approach is built upon a novel watermark definition that leverages patch-level luminance statistics. Specifically, the service provider predefines a binary pattern together with...

💬 0 commentsarXiv:2601.01085v1PDF
0

Posted in cs.CV · 2026-01-03 · Adari Rama Sukanya, Puvvula Roopesh Naga Sri Sai, Bodduru Neshika, Rimalapudi Sarvendranath

A UAV-Based Multispectral and RGB Dataset for Multi-Stage Paddy Crop Monitoring in Indian Agricultural Fields

We present a large-scale unmanned aerial vehicle (UAV)-based RGB and multispectral image dataset collected over paddy fields in the Vijayawada region, Andhra Pradesh, India, covering nursery to harvesting stages. We used a 20-megapixel RGB camera and a 5-megapixel four-band multispectral camera capturing red, green, red-edge, and...

💬 0 commentsarXiv:2601.01084v2PDF
0

Posted in cs.LG · 2026-01-03 · Bryon Tjanaka, Henry Chen, Matthew C. Fontaine, Stefanos Nikolaidis

Discount Model Search for Quality Diversity Optimization in High-Dimensional Measure Spaces

Quality diversity (QD) optimization searches for a collection of solutions that optimize an objective while attaining diverse outputs of a user-specified, vector-valued measure function. Contemporary QD algorithms are typically limited to low-dimensional measures because high-dimensional measures are prone to distortion, where many...

💬 0 commentsarXiv:2601.01082v5PDF
0

Posted in cs.IT · 2026-01-03 · Leilei Yu, Yunghsiang S. Han, Pingping Li, Jiasheng Yuan

A Novel Formula for Solving Quadratic Equations over Binary Extension Fields

Solving quadratic equations over finite fields is a fundamental task in algebraic coding theory and serves as a key subroutine for computing the roots of cubic and quartic polynomials. Notably, any quadratic polynomial over binary extension fields can be transformed into the reduced form $x^2+x+c\in \mathbb{F}_{2^m}[x]$, for which...

💬 0 commentsarXiv:2601.01079v2PDF
0

Posted in cs.CY · 2026-01-03 · Yuyan Huang, Haoran Li, Yifan Lu, Ruolin Wu, Siqian Chen, Chao Liu

Evidence-Grounded Multi-Agent Planning Support for Urban Carbon Governance via RAG

Urban carbon governance requires planners to integrate heterogeneous evidence -- emission inventories, statistical yearbooks, policy texts, technical measures, and academic findings -- into actionable, cross-departmental plans. Large Language Models (LLMs) can assist planning workflows, yet their factual reliability and evidential...

💬 0 commentsarXiv:2601.11587v1PDF
0

Posted in cs.LG · 2026-01-03 · Tanvi Jois, Hussain Ahmad, Fatima Noor, Faheem Ullah

Australian Bushfire Intelligence with AI-Driven Environmental Analytics

Bushfires are among the most destructive natural hazards in Australia, causing significant ecological, economic, and social damage. Accurate prediction of bushfire intensity is therefore essential for effective disaster preparedness and response. This study examines the predictive capability of spatio-temporal environmental data for...

💬 0 commentsarXiv:2601.06105v1PDF
0

Posted in cs.AI · 2026-01-03 · Krzysztof Sienicki

Comment on arXiv:2511.21731v1: Identifying Quantum Structure in AI Language: Evidence for Evolutionary Convergence of Human and Artificial Cognition

This note is a friendly technical check of arXiv:2511.21731v1. I highlight a few places where the manuscript's interpretation of (i) the reported CHSH/Bell-type calculations and (ii) Bose--Einstein (BE) fits to rank-frequency data seems to go beyond what the stated procedures can firmly support. I also point out one internal...

💬 0 commentsarXiv:2601.06104v1PDF
0

Posted in cs.LG · 2026-01-03 · Hansen Jin Lillemark, Benhao Huang, Fangneng Zhan, Yilun Du, Thomas Anderson Keller

Flow Equivariant World Models: Memory for Partially Observed Dynamic Environments

Embodied systems experience the world as 'a symphony of flows': a combination of many continuous streams of sensory input coupled to self-motion, interwoven with the dynamics of external objects. These sensory streams and the underlying dynamics of the world obey smooth, time-parameterized symmetries which existing world models...

💬 0 commentsarXiv:2601.01075v2PDF
0

Posted in cs.SI · 2026-01-03 · Erica Coppolillo, Emilio Ferrara

Gendered Pathways in AI Companionship: Cross-Community Behavior and Toxicity Patterns on Reddit

AI-companionship platforms are rapidly reshaping how people form emotional, romantic, and parasocial bonds with non-human agents, raising new questions about how these relationships intersect with gendered online behavior and exposure to harmful content. Focusing on the MyBoyfriendIsAI (MBIA) subreddit, we reconstruct the Reddit...

💬 0 commentsarXiv:2601.01073v1PDF
0

Posted in cs.LG · 2026-01-03 · Jing Wang, Peng Zhao, Zhi-Hua Zhou

Revisiting Weighted Strategy for Non-stationary Parametric Bandits and MDPs

Non-stationary parametric bandits have attracted much attention recently. There are three principled ways to deal with non-stationarity, including sliding-window, weighted, and restart strategies. As many non-stationary environments exhibit gradual drifting patterns, the weighted strategy is commonly adopted in real-world...

💬 0 commentsarXiv:2601.01069v1PDF
0

Posted in cs.CR · 2026-01-03 · Abdullah Al Mamun, Akid Abrar, Mizanur Rahman, M Sabbir Salek, Mashrur Chowdhury

Post-Quantum Cryptography for Intelligent Transportation Systems: An Implementation-Focused Review

As quantum computing advances, the cryptographic algorithms that underpin confidentiality, integrity, and authentication in Intelligent Transportation Systems (ITS) face increasing vulnerability to quantum-enabled attacks. To address these risks, governments and industry stakeholders are turning toward post-quantum cryptography (PQC),...

💬 0 commentsarXiv:2601.01068v2PDF
0

Posted in cs.RO · 2026-01-03 · Wenzheng Zhang, Yoshitaka Hara, Sousuke Nakamura

Topological Mapping and Navigation using a Monocular Camera based on AnyLoc

This paper proposes a method for topological mapping and navigation using a monocular camera. Based on AnyLoc, keyframes are converted into descriptors to construct topological relationships, enabling loop detection and map building. Unlike metric maps, topological maps simplify path planning and navigation by representing...

💬 0 commentsarXiv:2601.01067v1PDF
0

Posted in cs.LG · 2026-01-03 · Achraf Hsain, Yahya Zaki, Othman Abaakil, Hibat-allah Bekkar, Yousra Chtouki

Tiny Machine Learning for Real-Time Aquaculture Monitoring: A Case Study in Morocco

Aquaculture, the farming of aquatic organisms, is a rapidly growing industry facing challenges such as water quality fluctuations, disease outbreaks, and inefficient feed management. Traditional monitoring methods often rely on manual labor and are time consuming, leading to potential delays in addressing issues. This paper proposes...

💬 0 commentsarXiv:2601.01065v1PDF
0

Posted in cs.CV · 2026-01-03 · Jianan Li, Wangcai Zhao, Tingfa Xu

Efficient Hyperspectral Image Reconstruction Using Lightweight Separate Spectral Transformers

Hyperspectral imaging (HSI) is essential across various disciplines for its capacity to capture rich spectral information. However, efficiently reconstructing hyperspectral images from compressive sensing measurements presents significant challenges. To tackle these, we adopt a divide-and-conquer strategy that capitalizes on the...

💬 0 commentsarXiv:2601.01064v1PDF
0

Posted in cs.LG · 2026-01-03 · Yunlin Zeng

SPoRC-VIST: A Benchmark for Evaluating Generative Natural Narrative in Vision-Language Models

Vision-Language Models (VLMs) have achieved remarkable success in descriptive tasks such as image captioning and visual question answering (VQA). However, their ability to generate engaging, long-form narratives -- specifically multi-speaker podcast dialogues -- remains under-explored and difficult to evaluate. Standard metrics like...

💬 0 commentsarXiv:2601.01062v1PDF