paper-with-me

Papers

Learning Monocular Depth in Dynamic Environment via Context-aware Temporal Attention

2023-05-12 · Zizhang Wu, Zhuozheng Li, Zhi-Gang Fan, Yunzhe Wu, Yuanzhu Gan, Jian Pu, Xianzhi Li

The monocular depth estimation task has recently revealed encouraging prospects, especially for the autonomous driving task. To tackle the ill-posed problem of 3D geometric reasoning from 2D monocular images, multi-frame monocular methods are developed to leverage the perspective correlation information from sequential temporal frames. However, moving objects such as cars and trains usually violate the static scene assumption, leading to feature inconsistency deviation and misaligned cost values, which would mislead the optimization algorithm. In this work, we present CTA-Depth, a Context-aware Temporal Attention guided network for multi-frame monocular Depth estimation. Specifically, we first apply a multi-level attention enhancement module to integrate multi-level image features to obtain an initial depth and pose estimation. Then the proposed CTA-Refiner is adopted to alternatively optimize the depth and pose. During the refinement process, context-aware temporal attention (CTA) is developed to capture the global temporal-context correlations to maintain the feature consistency and estimation integrity of moving objects. In particular, we propose a long-range geometry embedding (LGE) module to produce a long-range temporal geometry prior. Our approach achieves significant improvements over state-of-the-art approaches on three benchmark datasets.

📄 PDF Abstract BibTeX arXiv:2305.07397

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous DrivingDepth EstimationMonocular Depth EstimationPose Estimation

Similar Papers 제목 키워드 기반

Depth-conditioned Dynamic Message Propagation for Monocular 3D Object Detection

2021-03-30 · CVPR 2021 1 · Li Wang, Liang Du, Xiaoqing Ye, Yanwei Fu 외

The objective of this paper is to learn context- and depth-aware feature representation to solve the problem of monocular 3D object detection. We make following contributions: (i) rather than appealing to the complicated…

3D Object DetectionMonocular 3D Object Detectionobject-detectionObject Detection

Disentangling Object Motion and Occlusion for Unsupervised Multi-frame Monocular Depth

2022-03-29 · Ziyue Feng, Liang Yang, Longlong Jing, HaiYan Wang 외

Conventional self-supervised monocular depth prediction methods are based on a static environment assumption, which leads to accuracy degradation in dynamic scenes due to the mismatch and occlusion problems introduced by…

Depth EstimationDepth PredictionDisentanglementMonocular Depth Estimation+4

No Pose Estimation? No Problem: Pose-Agnostic and Instance-Aware Test-Time Adaptation for Monocular Depth Estimation

2025-11-07 · Mingyu Sung, Hyeonmin Choe, Il-Min Kim, Sangseok Yun 외 arxiv

Monocular depth estimation (MDE), inferring pixel-level depths in single RGB images from a monocular camera, plays a crucial and pivotal role in a variety of AI applications demanding a three-dimensional (3D) topographic…

Monocular Depth EstimationPanoptic SegmentationTest-time AdaptationPose Estimation

MonoDTR: Monocular 3D Object Detection with Depth-Aware Transformer

2022-03-21 · CVPR 2022 1 · Kuan-Chih Huang, Tsung-Han Wu, Hung-Ting Su, Winston H. Hsu

Monocular 3D object detection is an important yet challenging task in autonomous driving. Some existing methods leverage depth information from an off-the-shelf depth estimator to assist 3D detection, but suffer from the…

3D Object Detection3D Object Detection From Monocular ImagesAutonomous DrivingMonocular 3D Object Detection+3

Manydepth2: Motion-Aware Self-Supervised Multi-Frame Monocular Depth Estimation in Dynamic Scenes

2023-12-23 · Kaichen Zhou, Jia-Wang Bian, Jian-Qing Zheng, JiaXing Zhong 외

Despite advancements in self-supervised monocular depth estimation, challenges persist in dynamic scenarios due to the dependence on assumptions about a static world. In this paper, we present Manydepth2, to achieve prec…

Camera Pose EstimationComputational EfficiencyDepth EstimationMonocular Depth Estimation+1