Decoupled Guidance Diffusion for Adaptive Offline Safe Reinforcement Learning
arXiv:2605.02777v2 Announce Type: replace-cross Abstract: Offline safe reinforcement learning often requires policies to adapt at deployment time to safety budgets…
arXiv:2605.02777v2 Announce Type: replace-cross Abstract: Offline safe reinforcement learning often requires policies to adapt at deployment time to safety budgets…
Nature Machine Intelligence, Published online: 07 May 2026; doi:10.1038/s42256-026-01235-7 Li and Walsh show that a unified ‘platonic’ geometry emerges across…
Nature Machine Intelligence, Published online: 07 May 2026; doi:10.1038/s42256-026-01241-9 A promising foundation model is developed for a range of downstream…
arXiv:2605.00382v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly deployed to generate code for human-centered applications where demographic…
arXiv:2304.14419v2 Announce Type: replace-cross Abstract: We propose a novel learning-based approach for robust 3D shape matching. Our method builds upon…
arXiv:2605.02011v1 Announce Type: cross Abstract: Automating the drafting of judgment documents is pivotal to judicial efficiency, yet it remains challenging…
arXiv:2605.02469v1 Announce Type: cross Abstract: Online reinforcement learning with verifiable rewards (RLVR) turns checkable outcomes into a scalable training signal,…
arXiv:2602.22480v2 Announce Type: replace Abstract: An important emerging application of coding agents is agent optimization: the iterative improvement of a…
arXiv:2605.02814v1 Announce Type: cross Abstract: Blind face restoration is highly ill-posed under severe degradation, where identity-critical details may be missing…