A Survey: Spatiotemporal Consistency in Video Generation

ByAdmin

Feb 19, 2026

THE AI TODAY

arXiv:2502.17863v2 Announce Type: replace-cross
Abstract: Video generation aims to produce temporally coherent sequences of visual frames, representing a pivotal advancement in Artificial Intelligence Generated Content (AIGC). Compared to static image generation, video generation poses unique challenges: it demands not only high-quality individual frames but also strong temporal coherence to ensure consistency throughout the spatiotemporal sequence. Although research addressing spatiotemporal consistency in video generation has increased in recent years, systematic reviews focusing on this core issue remain relatively scarce. To fill this gap, this paper views the video generation task as a sequential sampling process from a high-dimensional spatiotemporal distribution, and further discusses spatiotemporal consistency. We provide a systematic review of the latest advancements in the field. The content spans multiple dimensions including generation models, feature representations, generation frameworks, post-processing techniques, training strategies, benchmarks and evaluation metrics, with a particular focus on the mechanisms and effectiveness of various methods in maintaining spatiotemporal consistency. Finally, this paper explores future research directions and potential challenges in this field, aiming to provide valuable insights for advancing video generation technology. The project link is https://github.com/Yin-Z-Y/A-Survey-Spatiotemporal-Consistency-in-Video-Generation.

By Admin

AI RESEARCH

Accurate prediction of ecDNA in interphase cancer cells using deep neural networks

Apr 11, 2026 Admin

AI RESEARCH

Using causal machine learning and real world data to improve dose response decision making for secukinumab in psoriatic arthritis

Apr 11, 2026 Admin

AI RESEARCH

LobePrior segments lung lobes on computed tomography images in the presence of severe abnormalities

Apr 10, 2026 Admin

A Survey: Spatiotemporal Consistency in Video Generation

ByAdmin

By Admin

Related Post

Accurate prediction of ecDNA in interphase cancer cells using deep neural networks

Using causal machine learning and real world data to improve dose response decision making for secukinumab in psoriatic arthritis

LobePrior segments lung lobes on computed tomography images in the presence of severe abnormalities

You missed

Using causal machine learning and real world data to improve dose response decision making for secukinumab in psoriatic arthritis

Accurate prediction of ecDNA in interphase cancer cells using deep neural networks

A lightweight machine learning approach for DDoS detection and classification

LobePrior segments lung lobes on computed tomography images in the presence of severe abnormalities