paper-with-me

홈 › Papers

R3D2: Realistic 3D Asset Insertion via Diffusion for Autonomous Driving Simulation

2025-06-09 · William Ljungbergh, Bernardo Taveira, Wenzhao Zheng, Adam Tonderski, Chensheng Peng, Fredrik Kahl, Christoffer Petersson, Michael Felsberg, Kurt Keutzer, Masayoshi Tomizuka, Wei Zhan

Validating autonomous driving (AD) systems requires diverse and safety-critical testing, making photorealistic virtual environments essential. Traditional simulation platforms, while controllable, are resource-intensive to scale and often suffer from a domain gap with real-world data. In contrast, neural reconstruction methods like 3D Gaussian Splatting (3DGS) offer a scalable solution for creating photorealistic digital twins of real-world driving scenes. However, they struggle with dynamic object manipulation and reusability as their per-scene optimization-based methodology tends to result in incomplete object models with integrated illumination effects. This paper introduces R3D2, a lightweight, one-step diffusion model designed to overcome these limitations and enable realistic insertion of complete 3D assets into existing scenes by generating plausible rendering effects-such as shadows and consistent lighting-in real time. This is achieved by training R3D2 on a novel dataset: 3DGS object assets are generated from in-the-wild AD data using an image-conditioned 3D generative model, and then synthetically placed into neural rendering-based virtual environments, allowing R3D2 to learn realistic integration. Quantitative and qualitative evaluations demonstrate that R3D2 significantly enhances the realism of inserted assets, enabling use-cases like text-to-3D asset insertion and cross-scene/dataset object transfer, allowing for true scalability in AD validation. To promote further research in scalable and realistic AD simulation, we will release our dataset and code, see https://research.zenseact.com/publications/R3D2/.

📄 PDF Abstract BibTeX arXiv:2506.07826

Code (0)

등록된 구현이 없습니다.

Tasks

3DGSAutonomous DrivingNeural RenderingObjectText to 3D

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

SCPainter: A Unified Framework for Realistic 3D Asset Insertion and Novel View Synthesis

2025-12-27 · Paul Dobre, Jackson Cooper, Xin Wang, Hongzhou Yang arxiv

3D Asset insertion and novel view synthesis (NVS) are key components for autonomous driving simulation, enhancing the diversity of training data. With better training data that is diverse and covers a wide range of situa…

Novel View SynthesisAutonomous DrivingPoint Clouds

SymDrive: Realistic and Controllable Driving Simulator via Symmetric Auto-regressive Online Restoration

2025-12-25 · Zhiyuan Liu, Daocheng Fu, Pinlong Cai, Lening Wang 외 arxiv

High-fidelity and controllable 3D simulation is essential for addressing the long-tail data scarcity in Autonomous Driving (AD), yet existing methods struggle to simultaneously achieve photorealistic rendering and intera…

Novel View SynthesisAutonomous Driving

DriveWeaver: Point-Conditioned Video Inpainting for Controllable Vehicle Insertion in Autonomous Driving Simulation

2026-06-30 · Junzhe Jiang, Zipei Ma, Zijie Pan, Li Zhang arxiv

A pivotal step in autonomous driving simulation involves inserting foreground vehicles with predefined trajectories into simulated scenes. This process enhances scene diversity and facilitates the creation of various cor…

Autonomous DrivingVideo InpaintingPoint Clouds

VQA-Diff: Exploiting VQA and Diffusion for Zero-Shot Image-to-3D Vehicle Asset Generation in Autonomous Driving

2024-07-09 · Yibo Liu, Zheyuan Yang, Guile Wu, Yuan Ren 외

Generating 3D vehicle assets from in-the-wild observations is crucial to autonomous driving. Existing image-to-3D methods cannot well address this problem because they learn generation merely from image RGB information w…

Autonomous DrivingImage to 3DLanguage ModellingLarge Language Model+4

Mirage: One-Step Video Diffusion for Photorealistic and Coherent Asset Editing in Driving Scenes

2025-12-30 · Shuyun Wang, Haiyang Sun, Bing Wang, Hangjun Ye 외 arxiv

Vision-centric autonomous driving systems rely on diverse and scalable training data to achieve robust performance. While video object editing offers a promising path for data augmentation, existing methods often struggl…

Autonomous DrivingData Augmentation