paper-with-me

홈 › Papers

X-Scene: Large-Scale Driving Scene Generation with High Fidelity and Flexible Controllability

2025-06-16 · Yu Yang, Alan Liang, Jianbiao Mei, Yukai Ma, Yong liu, Gim Hee Lee

Diffusion models are advancing autonomous driving by enabling realistic data synthesis, predictive end-to-end planning, and closed-loop simulation, with a primary focus on temporally consistent generation. However, the generation of large-scale 3D scenes that require spatial coherence remains underexplored. In this paper, we propose X-Scene, a novel framework for large-scale driving scene generation that achieves both geometric intricacy and appearance fidelity, while offering flexible controllability. Specifically, X-Scene supports multi-granular control, including low-level conditions such as user-provided or text-driven layout for detailed scene composition and high-level semantic guidance such as user-intent and LLM-enriched text prompts for efficient customization. To enhance geometrical and visual fidelity, we introduce a unified pipeline that sequentially generates 3D semantic occupancy and the corresponding multiview images, while ensuring alignment between modalities. Additionally, we extend the generated local region into a large-scale scene through consistency-aware scene outpainting, which extrapolates new occupancy and images conditioned on the previously generated area, enhancing spatial continuity and preserving visual coherence. The resulting scenes are lifted into high-quality 3DGS representations, supporting diverse applications such as scene exploration. Comprehensive experiments demonstrate that X-Scene significantly advances controllability and fidelity for large-scale driving scene generation, empowering data generation and simulation for autonomous driving.

📄 PDF Abstract BibTeX arXiv:2506.13558

Code (0)

등록된 구현이 없습니다.

Tasks

3DGSAutonomous DrivingScene Generation

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

LSD-3D: Large-Scale 3D Driving Scene Generation with Geometry Grounding

2025-08-26 · Julian Ost, Andrea Ramazzina, Amogh Joshi, Maximilian Bömer 외 arxiv

Large-scale scene data is essential for training and testing in robot learning. Neural reconstruction methods have promised the capability of reconstructing large physically-grounded outdoor scenes from captured sensor d…

Novel View SynthesisScene Generation

SEM-ROVER: Semantic Voxel-Guided Diffusion for Large-Scale Driving Scene Generation

2026-04-07 · Hiba Dahmani, Nathan Piasco, Moussab Bennehar, Luis Roldão 외 arxiv

Scalable generation of outdoor driving scenes requires 3D representations that remain consistent across multiple viewpoints and scale to large areas. Existing solutions either rely on image or video generative models dis…

Scene Generation

VMA: Divide-and-Conquer Vectorized Map Annotation System for Large-Scale Driving Scene

2023-04-19 · Shaoyu Chen, Yunchi Zhang, Bencheng Liao, Jiafeng Xie 외

High-definition (HD) map serves as the essential infrastructure of autonomous driving. In this work, we build up a systematic vectorized map annotation framework (termed VMA) for efficiently generating HD map of large-sc…

Autonomous Driving

FreeGen: Feed-Forward Reconstruction-Generation Co-Training for Free-Viewpoint Driving Scene Synthesis

2025-12-04 · Shijie Chen, Peixi Peng arxiv

Closed-loop simulation and scalable pre-training for autonomous driving require synthesizing free-viewpoint driving scenes. However, existing datasets and generative pipelines rarely provide consistent off-trajectory obs…

Autonomous Driving

DriveGen3D: Boosting Feed-Forward Driving Scene Generation with Efficient Video Diffusion

2025-10-17 · Weijie Wang, Jiagang Zhu, Zeyu Zhang, Xiaofeng Wang 외 arxiv

We present DriveGen3D, a novel framework for generating high-quality and highly controllable dynamic 3D driving scenes that addresses critical limitations in existing methodologies. Current approaches to driving scene sy…

Scene GenerationVideo Generation