paper-with-me

홈 › Papers

Physical Informed Driving World Model

2024-12-11 · Zhuoran Yang, Xi Guo, Chenjing Ding, Chiyu Wang, Wei Wu

Autonomous driving requires robust perception models trained on high-quality, large-scale multi-view driving videos for tasks like 3D object detection, segmentation and trajectory prediction. While world models provide a cost-effective solution for generating realistic driving videos, challenges remain in ensuring these videos adhere to fundamental physical principles, such as relative and absolute motion, spatial relationship like occlusion and spatial consistency, and temporal consistency. To address these, we propose DrivePhysica, an innovative model designed to generate realistic multi-view driving videos that accurately adhere to essential physical principles through three key advancements: (1) a Coordinate System Aligner module that integrates relative and absolute motion features to enhance motion interpretation, (2) an Instance Flow Guidance module that ensures precise temporal consistency via efficient 3D flow extraction, and (3) a Box Coordinate Guidance module that improves spatial relationship understanding and accurately resolves occlusion hierarchies. Grounded in physical principles, we achieve state-of-the-art performance in driving video generation quality (3.96 FID and 38.06 FVD on the Nuscenes dataset) and downstream perception tasks. Our project homepage: https://metadrivescape.github.io/papers_project/DrivePhysica/page.html

📄 PDF Abstract BibTeX arXiv:2412.08410

Code (0)

등록된 구현이 없습니다.

Tasks

3D Object DetectionAutonomous Drivingmodelobject-detectionObject DetectionTrajectory PredictionVideo Generation

Similar Papers 제목 키워드 기반

Morpheus: Benchmarking Physical Reasoning of Video Generative Models with Real Physical Experiments

2025-04-03 · Chenyu Zhang, Daniil Cherniavskii, Andrii Zadaianchuk, Antonios Tragoudaras 외

Recent advances in image and video generation raise hopes that these models possess world modeling capabilities, the ability to generate realistic, physically plausible videos. This could revolutionize applications in ro…

Physical Commonsense ReasoningVideo Generation

Physics-informed Diffusion Mamba Transformer for Real-world Driving

2026-01-31 · Hang Zhou, Qiang Zhang, Peiran Liu, Yihao Qin 외 arxiv

Autonomous driving systems demand trajectory planners that not only model the inherent uncertainty of future motions but also respect complex temporal dependencies and underlying physical laws. While diffusion-based gene…

Autonomous DrivingMotion Planning

Distill to Think, Foresee to Act: Cognitive-Physical Reinforcement Learning for Autonomous Driving

2026-05-20 · Yang Wu, Qiang Meng, Zhaojiang Liu, Youquan Liu 외 arxiv

Current end-to-end autonomous driving models are fundamentally constrained by the behavioral cloning ceiling of imitation learning. While reinforcement learning offers a path to smarter autonomy, it demands two missing p…

Reinforcement LearningAutonomous Driving

A Causal Probabilistic Framework for Perception-Informed Closed-Loop Simulation of Autonomous Driving

2026-06-05 · Zhennan Fei, Rickard Johansson, Mikael Andersson, Matthias Eng 외 arxiv

Software-in-the-loop (SIL) simulation is a cornerstone for the validation of modern automotive safety functions. However, many current frameworks utilize ideal sensing, which bypasses the functional insufficiencies of pe…

Autonomous Driving

Risk-Controllable Multi-View Diffusion for Driving Scenario Generation

2026-03-12 · Hongyi Lin, Wenxiu Shi, Heye Huang, Dingyi Zhuang 외 arxiv

Generating safety-critical driving scenarios is crucial for evaluating and improving autonomous driving systems, but long-tail risky situations are rarely observed in real-world data and difficult to specify through manu…

Autonomous Driving