VLRS-Bench: A Vision-Language Reasoning Benchmark for Remote Sensing
arXiv:2602.07045v2 Announce Type: replace-cross Abstract: Recent advancements in Multimodal Large Language Models (MLLMs) have enabled complex reasoning. However, existing remote…
arXiv:2602.07045v2 Announce Type: replace-cross Abstract: Recent advancements in Multimodal Large Language Models (MLLMs) have enabled complex reasoning. However, existing remote…
arXiv:2605.13848v1 Announce Type: new Abstract: Agentic LLM frameworks that rely on prompted orchestration, where the model itself determines workflow transitions,…
arXiv:2605.13789v2 Announce Type: replace-cross Abstract: Protein structure tokenizers (PSTs) are workhorses in protein language modeling, function prediction, and evolutionary analysis.…
arXiv:2605.14712v1 Announce Type: cross Abstract: Robot imitation data are often multimodal: similar visual-language observations may be followed by different action…
arXiv:2605.14710v1 Announce Type: cross Abstract: Deep learning and multi-modal fusion have demonstrated transformative potential in medical diagnosis by integrating diverse…
arXiv:2605.13369v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are typically deployed with fixed parameters, and their performance is often…
Nature Machine Intelligence, Published online: 15 May 2026; doi:10.1038/s42256-026-01240-w A strong sustainability approach to AI development