paper-with-me

Papers

Evolution of Video Generative Foundations

2026-04-07 · Teng Hu, Jiangning Zhang, Hongrui Huang, Ran Yi, Zihan Su, Jieyu Weng, Zhucun Xue, Lizhuang Ma, Ming-Hsuan Yang, Dacheng Tao arxiv

The rapid advancement of Artificial Intelligence Generated Content (AIGC) has revolutionized video generation, enabling systems ranging from proprietary pioneers like OpenAI's Sora, Google's Veo3, and Bytedance's Seedance to powerful open-source contenders like Wan and HunyuanVideo to synthesize temporally coherent and semantically rich videos. These advancements pave the way for building "world models" that simulate real-world dynamics, with applications spanning entertainment, education, and virtual reality. However, existing reviews on video generation often focus on narrow technical fields, e.g., Generative Adversarial Networks (GAN) and diffusion models, or specific tasks (e. g., video editing), lacking a comprehensive perspective on the field's evolution, especially regarding Auto-Regressive (AR) models and integration of multimodal information. To address these gaps, this survey firstly provides a systematic review of the development of video generation technology, tracing its evolution from early GANs to dominant diffusion models, and further to emerging AR-based and multimodal techniques. We conduct an in-depth analysis of the foundational principles, key advancements, and comparative strengths/limitations. Then, we explore emerging trends in multimodal video generation, emphasizing the integration of diverse data types to enhance contextual awareness. Finally, by bridging historical developments and contemporary innovations, this survey offers insights to guide future research in video generation and its applications, including virtual/augmented reality, personalized education, autonomous driving simulations, digital entertainment, and advanced world models, in this rapidly evolving field. For more details, please refer to the project at https://github.com/sjtuplayer/Awesome-Video-Foundations.

📄 PDF Abstract BibTeX arXiv:2604.06339

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous DrivingVideo Generation

Similar Papers 제목 키워드 기반

Survey of Video Diffusion Models: Foundations, Implementations, and Applications

2025-04-22 · Yimu Wang, Xuye Liu, Wei Pang, Li Ma 외

Recent advances in diffusion models have revolutionized video generation, offering superior temporal consistency and visual quality compared to traditional generative adversarial networks-based approaches. While this eme…

Computational EfficiencyDenoisingQuestion AnsweringSuper-Resolution+2

Integrating Reinforcement Learning with Visual Generative Models: Foundations and Advances

2025-08-14 · Yuanzhi Liang, Yijie Fang, Ke Hao, Rui Li 외 arxiv

Generative models have made significant progress in synthesizing visual content, including images, videos, and 3D/4D structures. However, they are typically trained with surrogate objectives such as likelihood or reconst…

Reinforcement Learning

On logic and generative AI

2024-09-22 · Yuri Gurevich, Andreas Blass

A hundred years ago, logic was almost synonymous with foundational studies. The ongoing AI revolution raises many deep foundational problems involving neuroscience, philosophy, computer science, and logic. The goal of th…

Philosophy

EvoGM: Learning to Merge LLMs via Evolutionary Generative Optimization

2026-05-28 · Tao Jiang, Xinmeng Yu, Chenhao Yi, Yiling Wu 외 arxiv

Evolutionary model merging provides a powerful framework for the automated, training-free composition of LLMs through parameter-space search. However, existing methods predominantly rely on stochastic, hand-crafted opera…

Generative Models for Synthetic Data: Transforming Data Mining in the GenAI Era

2025-08-27 · Dawei Li, Yue Huang, Ming Li, Tianyi Zhou 외 arxiv

Generative models such as Large Language Models, Diffusion Models, and generative adversarial networks have recently revolutionized the creation of synthetic data, offering scalable solutions to data scarcity, privacy, a…

Synthetic Data Generation