TWLA: Achieving Ternary Weights and Low-Bit Activations for LLMs via Post-Training Quantization
arXiv:2606.13054v2 Announce Type: replace-cross Abstract: Large language models (LLMs) exhibit exceptional general language processing capabilities, but their memory and compute…
Standardized Methods and Recommendations for Green Federated Learning
arXiv:2602.00343v2 Announce Type: replace-cross Abstract: Federated learning (FL) enables collaborative model training over privacy-sensitive, distributed data, but its environmental impact…
Select and Improve: Understanding the Mechanics of Post-Training for Reasoning
arXiv:2606.13125v1 Announce Type: cross Abstract: Reinforcement learning has rapidly emerged as a key component in the training of reasoning and…
The Safety-Aware Denoiser for Text Diffusion Models
arXiv:2605.08116v2 Announce Type: replace-cross Abstract: Recent work on text diffusion models offers a promising alternative to autoregressive generation, but controlling…
Understanding the Rejection of Fixes Generated by Agentic Pull Requests — Insights from the AIDev Dataset
arXiv:2606.13468v1 Announce Type: cross Abstract: AI coding agents are increasingly used to generate pull requests (PRs) that propose code fixes…
Once-for-All: Scalable Simultaneous Forecasting via Equilibrium State Estimation
arXiv:2606.13285v1 Announce Type: cross Abstract: We introduce Equilibrium State Estimation (ESE), a novel paradigm for simultaneous prediction, where multiple interacting…
ToolSense: A Diagnostic Framework for Auditing Parametric Tool Knowledge in LLMs
arXiv:2606.12451v1 Announce Type: new Abstract: Large language models deployed as agents over large tool catalogs face a critical tool-retrieval bottleneck.…
Frozen Multimodal Embeddings for AI-Assisted Interview Assessment of Personality and Cognitive Ability
arXiv:2606.11930v2 Announce Type: replace-cross Abstract: Predicting psychological traits from asynchronous video interviews (AVIs) is a challenging problem in AI-assisted interview…
