VeRO: An Evaluation Harness for Agents to Optimize Agents
arXiv:2602.22480v2 Announce Type: replace Abstract: An important emerging application of coding agents is agent optimization: the iterative improvement of a…
arXiv:2602.22480v2 Announce Type: replace Abstract: An important emerging application of coding agents is agent optimization: the iterative improvement of a…
arXiv:2605.02814v1 Announce Type: cross Abstract: Blind face restoration is highly ill-posed under severe degradation, where identity-critical details may be missing…
arXiv:2605.02469v1 Announce Type: cross Abstract: Online reinforcement learning with verifiable rewards (RLVR) turns checkable outcomes into a scalable training signal,…
arXiv:2605.02124v1 Announce Type: cross Abstract: Softmax-routed mixture-of-experts models approach hard routing as the temperature tends to zero, but this limit…
Nature Machine Intelligence, Published online: 06 May 2026; doi:10.1038/s42256-026-01236-6 He et al. comprehensively test the reusability of PanPep, a meta-learning…
arXiv:2605.00650v1 Announce Type: cross Abstract: Fine-tuning LLMs is necessary for various dedicated downstream tasks, but classic backpropagation-based fine-tuning methods require…
arXiv:2509.24276v4 Announce Type: replace Abstract: Large language models (LLMs) excel at complex reasoning but remain limited by static and incomplete…
arXiv:2505.22003v2 Announce Type: replace-cross Abstract: In India, access to legal assistance for the general public has been observed to have…
arXiv:2510.15949v4 Announce Type: replace-cross Abstract: Large language models show promise for financial decision-making, yet deploying them as autonomous trading agents…
arXiv:2605.00654v1 Announce Type: cross Abstract: For a risk-averse finite-horizon Markov Decision Problem, we introduce a special class of Markov coherent…