paper-with-me

홈 › Papers

Hybrid Rendering for Multimodal Autonomous Driving: Merging Neural and Physics-Based Simulation

2025-03-12 · Máté Tóth, Péter Kovács, Zoltán Bendefy, Zoltán Hortsin, Balázs Teréki, Tamás Matuszka

Neural reconstruction models for autonomous driving simulation have made significant strides in recent years, with dynamic models becoming increasingly prevalent. However, these models are typically limited to handling in-domain objects closely following their original trajectories. We introduce a hybrid approach that combines the strengths of neural reconstruction with physics-based rendering. This method enables the virtual placement of traditional mesh-based dynamic agents at arbitrary locations, adjustments to environmental conditions, and rendering from novel camera viewpoints. Our approach significantly enhances novel view synthesis quality -- especially for road surfaces and lane markings -- while maintaining interactive frame rates through our novel training method, NeRF2GS. This technique leverages the superior generalization capabilities of NeRF-based methods and the real-time rendering speed of 3D Gaussian Splatting (3DGS). We achieve this by training a customized NeRF model on the original images with depth regularization derived from a noisy LiDAR point cloud, then using it as a teacher model for 3DGS training. This process ensures accurate depth, surface normals, and camera appearance modeling as supervision. With our block-based training parallelization, the method can handle large-scale reconstructions (greater than or equal to 100,000 square meters) and predict segmentation masks, surface normals, and depth maps. During simulation, it supports a rasterization-based rendering backend with depth-based composition and multiple camera models for real-time camera simulation, as well as a ray-traced backend for precise LiDAR simulation.

📄 PDF Abstract BibTeX arXiv:2503.09464

Code (0)

등록된 구현이 없습니다.

Tasks

3DGSAutonomous DrivingNeRFNovel View Synthesis

Methods 이 논문이 사용한 방법론

SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…

Similar Papers 제목 키워드 기반

Lightning NeRF: Efficient Hybrid Scene Representation for Autonomous Driving

2024-03-09 · Junyi Cao, Zhichao Li, Naiyan Wang, Chao Ma

Recent studies have highlighted the promising application of NeRF in autonomous driving contexts. However, the complexity of outdoor environments, combined with the restricted viewpoints in driving scenarios, complicates…

Autonomous DrivingNeRFNovel View Synthesis

Large Multimodal Models for Embodied Intelligent Driving: The Next Frontier in Self-Driving?

2026-01-13 · Long Zhang, Yuchen Xia, Bingqing Wei, Zhen Liu 외 arxiv

The advent of Large Multimodal Models (LMMs) offers a promising technology to tackle the limitations of modular design in autonomous driving, which often falters in open-world scenarios requiring sustained environmental …

Reinforcement LearningAutonomous DrivingLogical Reasoning

Towards Knowledge-driven Autonomous Driving

2023-12-07 · Xin Li, Yeqi Bai, Pinlong Cai, Licheng Wen 외

This paper explores the emerging knowledge-driven autonomous driving technologies. Our investigation highlights the limitations of current autonomous driving systems, in particular their sensitivity to data bias, difficu…

Autonomous DrivingNeural Rendering

Learn-to-Race: A Multimodal Control Environment for Autonomous Racing

2021-03-22 · ICCV 2021 10 · James Herman, Jonathan Francis, Siddha Ganju, Bingqing Chen 외

Existing research on autonomous driving primarily focuses on urban driving, which is insufficient for characterising the complex driving behaviour underlying high-speed racing. At the same time, existing racing simulatio…

Autonomous DrivingAutonomous RacingTrajectory Prediction

BEVWorld: A Multimodal World Model for Autonomous Driving via Unified BEV Latent Space

2024-07-08 · Yumeng Zhang, Shi Gong, Kaixin Xiong, Xiaoqing Ye 외

World models are receiving increasing attention in autonomous driving for their ability to predict potential future scenarios. In this paper, we present BEVWorld, a novel approach that tokenizes multimodal sensor inputs …

Autonomous DrivingDecodermotion prediction