paper-with-me

홈 › Papers

On depth prediction for autonomous driving using self-supervised learning

2024-03-10 · Houssem Boulahbal

Perception of the environment is a critical component for enabling autonomous driving. It provides the vehicle with the ability to comprehend its surroundings and make informed decisions. Depth prediction plays a pivotal role in this process, as it helps the understanding of the geometry and motion of the environment. This thesis focuses on the challenge of depth prediction using monocular self-supervised learning techniques. The problem is approached from a broader perspective first, exploring conditional generative adversarial networks (cGANs) as a potential technique to achieve better generalization was performed. In doing so, a fundamental contribution to the conditional GANs, the acontrario cGAN was proposed. The second contribution entails a single image-to-depth self-supervised method, proposing a solution for the rigid-scene assumption using a novel transformer-based method that outputs a pose for each dynamic object. The third significant aspect involves the introduction of a video-to-depth map forecasting approach. This method serves as an extension of self-supervised techniques to predict future depths. This involves the creation of a novel transformer model capable of predicting the future depth of a given scene. Moreover, the various limitations of the aforementioned methods were addressed and a video-to-video depth maps model was proposed. This model leverages the spatio-temporal consistency of the input and output sequence to predict a more accurate depth sequence output. These methods have significant applications in autonomous driving (AD) and advanced driver assistance systems (ADAS).

📄 PDF Abstract BibTeX arXiv:2403.06194

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous DrivingDepth EstimationDepth PredictionSelf-Supervised Learning

Similar Papers 제목 키워드 기반

FSNet: Redesign Self-Supervised MonoDepth for Full-Scale Depth Prediction for Autonomous Driving

2023-04-21 · Yuxuan Liu, Zhenhua Xu, Huaiyang Huang, Lujia Wang 외

Predicting accurate depth with monocular images is important for low-cost robotic applications and autonomous driving. This study proposes a comprehensive self-supervised framework for accurate scale-aware depth predicti…

Autonomous DrivingDepth EstimationDepth PredictionOptical Flow Estimation+2

Diffusion-Augmented Depth Prediction with Sparse Annotations

2023-08-04 · Jiaqi Li, Yiran Wang, Zihao Huang, Jinghong Zheng 외

Depth estimation aims to predict dense depth maps. In autonomous driving scenes, sparsity of annotations makes the task challenging. Supervised models produce concave objects due to insufficient structural information. T…

Autonomous DrivingDepth EstimationDepth PredictionPose Estimation+2

PlaneDepth: Self-supervised Depth Estimation via Orthogonal Planes

2022-10-04 · CVPR 2023 1 · Ruoyu Wang, Zehao Yu, Shenghua Gao

Multiple near frontal-parallel planes based depth representation demonstrated impressive results in self-supervised monocular depth estimation (MDE). Whereas, such a representation would cause the discontinuity of the gr…

Autonomous DrivingData AugmentationDepth EstimationMonocular Depth Estimation

MGNiceNet: Unified Monocular Geometric Scene Understanding

2024-11-18 · Markus Schön, Michael Buchholz, Klaus Dietmayer

Monocular geometric scene understanding combines panoptic segmentation and self-supervised depth estimation, focusing on real-time application in autonomous vehicles. We introduce MGNiceNet, a unified approach that uses …

Autonomous DrivingAutonomous VehiclesDepth EstimationDepth Prediction+5

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving

2024-11-22 · CVPR 2025 1 · Haiming Zhang, Wending Zhou, Yiyao Zhu, Xu Yan 외

This paper introduces VisionPAD, a novel self-supervised pre-training paradigm designed for vision-centric algorithms in autonomous driving. In contrast to previous approaches that employ neural rendering with explicit d…

3D Object DetectionAutonomous DrivingNeural Renderingobject-detection+1