Metric-Gradient Projection for Stable Multi-Agent Policy Learning
arXiv:2605.18809v1 Announce Type: cross Abstract: General-sum multi-agent learning is often governed by a stacked update field in which each agent’s…
arXiv:2605.18809v1 Announce Type: cross Abstract: General-sum multi-agent learning is often governed by a stacked update field in which each agent’s…
arXiv:2605.16447v2 Announce Type: replace-cross Abstract: Spatiotemporal forecasting is critical for real-world applications like traffic management, yet capturing reliable interactions remains…
arXiv:2605.15652v2 Announce Type: replace-cross Abstract: Vector-HaSH and the Tolman-Eichenbaum Machine propose the hippocampal-entorhinal circuit factorizes content from a grid-cell scaffold,…
arXiv:2605.17937v1 Announce Type: cross Abstract: Quantitative backtesting is essential for evaluating trading strategies but remains hampered by high technical barriers…
arXiv:2605.17938v1 Announce Type: cross Abstract: Training data attribution (TDA) should enable generative model interpretability and foster a variety of related…
arXiv:2605.16234v2 Announce Type: replace-cross Abstract: When researchers ask whether two transformer layers are “equivalent” for compression, they often conflate distinct…
arXiv:2605.16265v1 Announce Type: new Abstract: The safety of autonomous AI agents is increasingly recognized as a critical open problem. As…
arXiv:2604.04202v2 Announce Type: replace-cross Abstract: AI agents deployed as persistent assistants must maintain correct beliefs as their information environment evolves.…
arXiv:2605.17839v1 Announce Type: cross Abstract: Knowledge distillation transfers knowledge from a high capacity teacher to a compact student using a…
arXiv:2604.09297v2 Announce Type: replace-cross Abstract: Agent skills are increasingly used to configure coding agents for software engineering (SE) tasks, yet…