paper-with-me

Papers

MagicDrive3D: Controllable 3D Generation for Any-View Rendering in Street Scenes

2024-05-23 · Ruiyuan Gao, Kai Chen, Zhihao LI, Lanqing Hong, Zhenguo Li, Qiang Xu

While controllable generative models for images and videos have achieved remarkable success, high-quality models for 3D scenes, particularly in unbounded scenarios like autonomous driving, remain underdeveloped due to high data acquisition costs. In this paper, we introduce MagicDrive3D, a novel pipeline for controllable 3D street scene generation that supports multi-condition control, including BEV maps, 3D objects, and text descriptions. Unlike previous methods that reconstruct before training the generative models, MagicDrive3D first trains a video generation model and then reconstructs from the generated data. This innovative approach enables easily controllable generation and static scene acquisition, resulting in high-quality scene reconstruction. To address the minor errors in generated content, we propose deformable Gaussian splatting with monocular depth initialization and appearance modeling to manage exposure discrepancies across viewpoints. Validated on the nuScenes dataset, MagicDrive3D generates diverse, high-quality 3D driving scenes that support any-view rendering and enhance downstream tasks like BEV segmentation. Our results demonstrate the framework's superior performance, showcasing its potential for autonomous driving simulation and beyond.

📄 PDF Abstract BibTeX arXiv:2405.14475

Code (0)

등록된 구현이 없습니다.

Tasks

3D GenerationAutonomous DrivingBEV SegmentationScene GenerationVideo Generation

Similar Papers 제목 키워드 기반

MagicDrive: Street View Generation with Diverse 3D Geometry Control

2023-10-04 · Ruiyuan Gao, Kai Chen, Enze Xie, Lanqing Hong 외

Recent advancements in diffusion models have significantly enhanced the data synthesis with 2D control. Yet, precise 3D control in street view generation, crucial for 3D perception tasks, remains elusive. Specifically, u…

3D geometry3D Object DetectionBEV SegmentationObject+2

MagicDriveDiT: High-Resolution Long Video Generation for Autonomous Driving with Adaptive Control

2024-11-21 · Ruiyuan Gao, Kai Chen, Bo Xiao, Lanqing Hong 외

The rapid advancement of diffusion models has greatly improved video synthesis, especially in controllable video generation, which is essential for applications like autonomous driving. However, existing methods are limi…

Autonomous DrivingVideo Generation

StreetCrafter: Street View Synthesis with Controllable Video Diffusion Models

2024-12-17 · CVPR 2025 1 · Yunzhi Yan, Zhen Xu, Haotong Lin, Haian Jin 외

This paper aims to tackle the problem of photorealistic view synthesis from vehicle sensor data. Recent advancements in neural scene representation have achieved notable success in rendering high-quality autonomous drivi…

Autonomous DrivingNovel View Synthesis

Text2Street: Controllable Text-to-image Generation for Street Views

2024-02-07 · Jinming Su, Songen Gu, Yiting Duan, Xingyue Chen 외

Text-to-image generation has made remarkable progress with the emergence of diffusion models. However, it is still a difficult task to generate images for street views based on text, mainly because the road topology of s…

Image GenerationLayout GenerationObjectText to Image Generation+1

PerLDiff: Controllable Street View Synthesis Using Perspective-Layout Diffusion Models

2024-07-08 · Jinhua Zhang, Hualian Sheng, Sijia Cai, Bing Deng 외

Controllable generation is considered a potentially vital approach to address the challenge of annotating 3D data, and the precision of such controllable generation becomes particularly imperative in the context of data …

Autonomous DrivingImage Generation