paper-with-me

홈 › Papers

SIMSplat: Language-Aligned 4D Gaussian Splatting for Driving Scenario Generation

2025-10-02 · Sung-Yeon Park, Adam Lee, Juanwu Lu, Can Cui, Luyang Jiang, Rohit Gupta, Kyungtae Han, Ahmadreza Moradipari, Ziran Wang arxiv

Driving scene manipulation using real-world sensor data has emerged as a promising alternative to traditional driving simulators. Despite advances in language control and neural scene representations, existing methods treat grounding, editing, and simulation as loosely connected stages, relying on heuristic object localization, manual guidance, and single-agent validation, thereby constraining semantic expressiveness and hindering scalable, reactive scenario generation. We introduce SIMSplat, a driving scene editor built on scene-graph-based 4D Gaussian Splatting augmented with language-aligned features. By embedding appearance, motion, and location semantics directly into Gaussian scene-graph nodes, SIMSplat makes reconstructed scenes queryable through free-form natural language, bridging language understanding to object-level editing and multi-agent simulation within a single framework. Building on this language-grounded scene graph, SIMSplat supports diverse edits including fine-grained pedestrian manipulation, while a multi-agent path refinement module propagates changes across all agents to ensure reactive, physically plausible simulations. The pipeline further integrates with Vision-Language Models for automated scenario mining. Experiments show that SIMSplat more than doubles baseline grounding accuracy, achieves the highest task completion rate, and produces the lowest failure rates across diverse driving scenarios.

📄 PDF Abstract BibTeX arXiv:2510.02469

Code (0)

등록된 구현이 없습니다.

Tasks

Object Localization

Similar Papers 제목 키워드 기반

LR-SGS: Robust LiDAR-Reflectance-Guided Salient Gaussian Splatting for Self-Driving Scene Reconstruction

2026-03-13 · ZY Chen, F Zhu, H Zhu, DY Kong 외 arxiv

Recent 3D Gaussian Splatting (3DGS) methods have demonstrated the feasibility of self-driving scene reconstruction and novel view synthesis. However, most existing methods either rely solely on cameras or use LiDAR only …

Novel View SynthesisPoint Clouds

PointForward: Feedforward Driving Reconstruction through Point-Aligned Representations

2026-05-12 · Cheng Chi, Xianqi Wang, Hongcheng Luo, Mingfei Tu 외 arxiv

High-fidelity reconstruction of driving scenes is crucial for autonomous driving. While recent feedforward 3D Gaussian Splatting (3DGS) methods enable fast reconstruction, their per-pixel Gaussian prediction paradigm oft…

Autonomous Driving

DrivingGaussian: Composite Gaussian Splatting for Surrounding Dynamic Autonomous Driving Scenes

2023-12-13 · CVPR 2024 1 · Xiaoyu Zhou, Zhiwei Lin, Xiaojun Shan, Yongtao Wang 외

We present DrivingGaussian, an efficient and effective framework for surrounding dynamic autonomous driving scenes. For complex scenes with moving objects, we first sequentially and progressively model the static backgro…

Autonomous Driving

LT-Gaussian: Long-Term Map Update Using 3D Gaussian Splatting for Autonomous Driving

2025-08-03 · Luqi Cheng, Zhangshuo Qi, Zijie Zhou, Chao Lu 외 arxiv

Maps play an important role in autonomous driving systems. The recently proposed 3D Gaussian Splatting (3D-GS) produces rendering-quality explicit scene reconstruction results, demonstrating the potential for map constru…

Autonomous DrivingChange Detection

GaussianVLM: Scene-centric 3D Vision-Language Models using Language-aligned Gaussian Splats for Embodied Reasoning and Beyond

2025-07-01 · Anna-Maria Halacheva, Jan-Nico Zaech, Xi Wang, Danda Pani Paudel 외 arxiv

As multimodal language models advance, their application to 3D scene understanding is a fast-growing frontier, driving the development of 3D Vision-Language Models (VLMs). Current methods show strong dependence on object…

Scene Understanding