paper-with-me

홈 › Papers

StereoCarla: A High-Fidelity Driving Dataset for Generalizable Stereo

2025-09-16 · Xianda Guo, Chenming Zhang, Ruilin Wang, Youmin Zhang, Wenzhao Zheng, Matteo Poggi, Hao Zhao, Qin Zou, Long Chen arxiv

Stereo matching plays a crucial role in enabling depth perception for autonomous driving and robotics. While recent years have witnessed remarkable progress in stereo matching algorithms, largely driven by learning-based methods and synthetic datasets, the generalization performance of these models remains constrained by the limited diversity of existing training data. To address these challenges, we present StereoCarla, a high-fidelity synthetic stereo dataset specifically designed for autonomous driving scenarios. Built on the CARLA simulator, StereoCarla incorporates a wide range of camera configurations, including diverse baselines, viewpoints, and sensor placements as well as varied environmental conditions such as lighting changes, weather effects, and road geometries. We conduct comprehensive cross-domain experiments across four standard evaluation datasets (KITTI2012, KITTI2015, Middlebury, ETH3D) and demonstrate that models trained on StereoCarla outperform those trained on 11 existing stereo datasets in terms of generalization accuracy across multiple benchmarks. Furthermore, when integrated into multi-dataset training, StereoCarla contributes substantial improvements to generalization accuracy, highlighting its compatibility and scalability. This dataset provides a valuable benchmark for developing and evaluating stereo algorithms under realistic, diverse, and controllable settings, facilitating more robust depth perception systems for autonomous vehicles. Code can be available at https://github.com/XiandaGuo/OpenStereo, and data can be available at https://xiandaguo.net/StereoCarla.

📄 PDF Abstract BibTeX arXiv:2509.12683

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous VehiclesAutonomous Driving

Similar Papers 제목 키워드 기반

Vista: A Generalizable Driving World Model with High Fidelity and Versatile Controllability

2024-05-27 · Shenyuan Gao, Jiazhi Yang, Li Chen, Kashyap Chitta 외

World models can foresee the outcomes of different actions, which is of paramount importance for autonomous driving. Nevertheless, existing driving world models still have limitations in generalization to unseen environm…

Autonomous DrivingVideo Generation

DreamDrive: Generative 4D Scene Modeling from Street View Images

2024-12-31 · Jiageng Mao, Boyi Li, Boris Ivanovic, Yuxiao Chen 외

Synthesizing photo-realistic visual observations from an ego vehicle's driving trajectory is a critical step towards scalable training of self-driving models. Reconstruction-based methods create 3D scenes from driving lo…

Autonomous DrivingNeural RenderingScene GenerationVideo Generation

MaskGWM: A Generalizable Driving World Model with Video Mask Reconstruction

2025-02-17 · CVPR 2025 1 · Jingcheng Ni, Yuxin Guo, Yichen Liu, Rui Chen 외

World models that forecast environmental changes from actions are vital for autonomous driving models with strong generalization. The prevailing driving world model mainly build on video prediction model. Although these …

2kAutonomous DrivingTask 2Video Prediction

Diffusion-guided Generalizable Enhancer for Urban Scene Reconstruction

2026-05-21 · Henry Che, Jingkang Wang, Yun Chen, Ze Yang 외 arxiv

Urban scene reconstruction from real-world observations has emerged as a powerful tool for self-driving development and testing. While current neural rendering approaches achieve high-fidelity rendering along the recorde…

Autonomous Driving

LokiTalk: Learning Fine-Grained and Generalizable Correspondences to Enhance NeRF-based Talking Head Synthesis

2024-11-29 · Tianqi Li, Ruobing Zheng, Bonan Li, ZiCheng Zhang 외

Despite significant progress in talking head synthesis since the introduction of Neural Radiance Fields (NeRF), visual artifacts and high training costs persist as major obstacles to large-scale commercial adoption. We p…

NeRFTransfer Learning