G-reasoner: Foundation Models for Unified Reasoning over Graph-structured Knowledge
arXiv:2509.24276v4 Announce Type: replace Abstract: Large language models (LLMs) excel at complex reasoning but remain limited by static and incomplete…
PORTool: Importance-Aware Policy Optimization with Rewarded Tree for Multi-Tool-Integrated Reasoning
arXiv:2510.26020v2 Announce Type: replace-cross Abstract: Multi-tool-integrated reasoning enables LLM-empowered tool-use agents to solve complex tasks by interleaving natural-language reasoning with…
How Well Does GPT-4o Understand Vision? Evaluating Multimodal Foundation Models on Standard Computer Vision Tasks
arXiv:2507.01955v3 Announce Type: replace-cross Abstract: Multimodal foundation models (MFMs), such as GPT-4o, have recently made remarkable progress. However, their detailed…
ATLAS: Adaptive Trading with LLM AgentS Through Dynamic Prompt Optimization and Multi-Agent Coordination
arXiv:2510.15949v4 Announce Type: replace-cross Abstract: Large language models show promise for financial decision-making, yet deploying them as autonomous trading agents…
Lightweight Domain Adaptation of a Large Language Model for Legal Assistance in the Indian Context
arXiv:2505.22003v2 Announce Type: replace-cross Abstract: In India, access to legal assistance for the general public has been observed to have…
TADI: Tool-Augmented Drilling Intelligence via Agentic LLM Orchestration over Heterogeneous Wellsite Data
arXiv:2605.00060v1 Announce Type: new Abstract: We present TADI (Tool-Augmented Drilling Intelligence), an agentic AI system that transforms drilling operational data…
Claw-Eval-Live: A Live Agent Benchmark for Evolving Real-World Workflows
arXiv:2604.28139v2 Announce Type: replace-cross Abstract: LLM agents are expected to complete end-to-end units of work across software tools, business services,…
Reinforcement Learning with Markov Risk Measures and Multipattern Risk Approximation
arXiv:2605.00654v1 Announce Type: cross Abstract: For a risk-averse finite-horizon Markov Decision Problem, we introduce a special class of Markov coherent…
AdaMeZO: Adam-style Zeroth-Order Optimizer for LLM Fine-tuning Without Maintaining the Moments
arXiv:2605.00650v1 Announce Type: cross Abstract: Fine-tuning LLMs is necessary for various dedicated downstream tasks, but classic backpropagation-based fine-tuning methods require…
AI Inference as Relocatable Electricity Demand: A Latency-Constrained Energy-Geography Framework
arXiv:2604.27855v2 Announce Type: replace-cross Abstract: AI inference is becoming a persistent and geographically distributed source of electricity demand. Unlike many…
