Reasoning Pattern Matters: Learning to Reason without Human Rationales
arXiv:2510.12643v1 Announce Type: cross Abstract: Large Language Models (LLMs) have demonstrated remarkable reasoning capabilities under the widely adopted SFT+RLVR paradigm,…
Aixel: A Unified, Adaptive and Extensible System for AI-powered Data Analysis
arXiv:2510.12642v1 Announce Type: cross Abstract: A growing trend in modern data analysis is the integration of data management with learning,…
AndesVL Technical Report: An Efficient Mobile-side Multimodal Large Language Model
arXiv:2510.11496v2 Announce Type: replace-cross Abstract: In recent years, while cloud-based MLLMs such as QwenVL, InternVL, GPT-4o, Gemini, and Claude Sonnet…
The Geometry of Reasoning: Flowing Logics in Representation Space
arXiv:2510.09782v1 Announce Type: new Abstract: We study how large language models (LLMs) “think” through their representation space. We propose a…
SPG: Sandwiched Policy Gradient for Masked Diffusion Language Models
arXiv:2510.09541v2 Announce Type: replace-cross Abstract: Diffusion large language models (dLLMs) are emerging as an efficient alternative to autoregressive models due…
video-SALMONN S: Streaming Audio-Visual LLMs Beyond Length Limits via Memory
arXiv:2510.11129v1 Announce Type: cross Abstract: Continuous, high-frame-rate, high-resolution processing of long video streams is critical for future AI agents, yet…
PhysioME: A Robust Multimodal Self-Supervised Framework for Physiological Signals with Missing Modalities
arXiv:2510.11110v1 Announce Type: cross Abstract: Missing or corrupted modalities are common in physiological signal-based medical applications owing to hardware constraints…
ChoirRec: Semantic User Grouping via LLMs for Conversion Rate Prediction of Low-Activity Users
arXiv:2510.09393v2 Announce Type: replace-cross Abstract: Accurately predicting conversion rates (CVR) for low-activity users remains a fundamental challenge in large-scale e-commerce…
Hypothesis Hunting with Evolving Networks of Autonomous Scientific Agents
arXiv:2510.08619v1 Announce Type: new Abstract: Large-scale scientific datasets — spanning health biobanks, cell atlases, Earth reanalyses, and more — create…
On the optimization dynamics of RLVR: Gradient gap and step size thresholds
arXiv:2510.08539v2 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR), which uses simple binary feedback to post-train large language…
