CanvasMAR: Improving Masked Autoregressive Video Generation With Canvas
arXiv:2510.13669v1 Announce Type: cross Abstract: Masked autoregressive models (MAR) have recently emerged as a powerful paradigm for image and video…
arXiv:2510.13669v1 Announce Type: cross Abstract: Masked autoregressive models (MAR) have recently emerged as a powerful paradigm for image and video…
arXiv:2510.11496v2 Announce Type: replace-cross Abstract: In recent years, while cloud-based MLLMs such as QwenVL, InternVL, GPT-4o, Gemini, and Claude Sonnet…
arXiv:2510.12642v1 Announce Type: cross Abstract: A growing trend in modern data analysis is the integration of data management with learning,…
arXiv:2510.12643v1 Announce Type: cross Abstract: Large Language Models (LLMs) have demonstrated remarkable reasoning capabilities under the widely adopted SFT+RLVR paradigm,…
arXiv:2510.11683v2 Announce Type: replace-cross Abstract: A key challenge in applying reinforcement learning (RL) to diffusion large language models (dLLMs) lies…
arXiv:2510.11736v1 Announce Type: new Abstract: This study evaluates Artificial Intelligence (AI) agents for Dhumbal, a culturally significant multiplayer card game…
arXiv:2510.09393v2 Announce Type: replace-cross Abstract: Accurately predicting conversion rates (CVR) for low-activity users remains a fundamental challenge in large-scale e-commerce…
arXiv:2510.11110v1 Announce Type: cross Abstract: Missing or corrupted modalities are common in physiological signal-based medical applications owing to hardware constraints…
arXiv:2510.11129v1 Announce Type: cross Abstract: Continuous, high-frame-rate, high-resolution processing of long video streams is critical for future AI agents, yet…
arXiv:2510.09541v2 Announce Type: replace-cross Abstract: Diffusion large language models (dLLMs) are emerging as an efficient alternative to autoregressive models due…