paper-with-me

Papers

MyGo: Consistent and Controllable Multi-View Driving Video Generation with Camera Control

2024-09-10 · Yining Yao, Xi Guo, Chenjing Ding, Wei Wu

High-quality driving video generation is crucial for providing training data for autonomous driving models. However, current generative models rarely focus on enhancing camera motion control under multi-view tasks, which is essential for driving video generation. Therefore, we propose MyGo, an end-to-end framework for video generation, introducing motion of onboard cameras as conditions to make progress in camera controllability and multi-view consistency. MyGo employs additional plug-in modules to inject camera parameters into the pre-trained video diffusion model, which retains the extensive knowledge of the pre-trained model as much as possible. Furthermore, we use epipolar constraints and neighbor view information during the generation process of each view to enhance spatial-temporal consistency. Experimental results show that MyGo has achieved state-of-the-art results in both general camera-controlled video generation and multi-view driving video generation tasks, which lays the foundation for more accurate environment simulation in autonomous driving. Project page: https://metadrivescape.github.io/papers_project/MyGo/page.html

📄 PDF Abstract BibTeX arXiv:2409.06189

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous DrivingVideo Generation

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…
Focus 설명 없음

Similar Papers 제목 키워드 기반

DrivingGaussian++: Towards Realistic Reconstruction and Editable Simulation for Surrounding Dynamic Driving Scenes

2025-08-28 · Yajiao Xiong, Xiaoyu Zhou, Yongtao Wan, Deqing Sun 외 arxiv

We present DrivingGaussian++, an efficient and effective framework for realistic reconstructing and controllable editing of surrounding dynamic autonomous driving scenes. DrivingGaussian++ models the static background us…

Autonomous Driving

BEV-VAE: Multi-view Image Generation with Spatial Consistency for Autonomous Driving

2025-07-01 · Zeming Chen, Hang Zhao arxiv

Multi-view image generation in autonomous driving demands consistent 3D scene understanding across camera views. Most existing methods treat this problem as a 2D image set generation task, lacking explicit 3D modeling. H…

Scene UnderstandingAutonomous DrivingScene GenerationImage Generation

V2VCrafter: Consistent Street-View Image Generation Across Vehicles

2026-05-28 · Yihang Tao, Yu Guo, Senkang Hu, Yanan Ma 외 arxiv

Connected and autonomous driving (CAD) systems leverage vehicle-to-vehicle (V2V) communication for multi-agent collaborative perception, yet remain constrained by scarce annotated real-world V2V datasets and limited gene…

3D Object DetectionData AugmentationImage Generation

Driving into the Future: Multiview Visual Forecasting and Planning with World Model for Autonomous Driving

2023-11-29 · CVPR 2024 1 · Yuqi Wang, JiaWei He, Lue Fan, Hongxin Li 외

In autonomous driving, predicting future events in advance and evaluating the foreseeable risks empowers autonomous vehicles to better plan their actions, enhancing safety and efficiency on the road. To this end, we prop…

Autonomous DrivingAutonomous Vehicles

CoGen: 3D Consistent Video Generation via Adaptive Conditioning for Autonomous Driving

2025-03-28 · Yishen Ji, Ziyue Zhu, Zhenxin Zhu, Kaixin Xiong 외

Recent progress in driving video generation has shown significant potential for enhancing self-driving systems by providing scalable and controllable training data. Although pretrained state-of-the-art generation models,…

3D GenerationAutonomous DrivingVideo Generation