paper-with-me

홈 › Papers

Virtually Enriched NYU Depth V2 Dataset for Monocular Depth Estimation: Do We Need Artificial Augmentation?

2024-04-15 · Dmitry Ignatov, Andrey Ignatov, Radu Timofte

We present ANYU, a new virtually augmented version of the NYU depth v2 dataset, designed for monocular depth estimation. In contrast to the well-known approach where full 3D scenes of a virtual world are utilized to generate artificial datasets, ANYU was created by incorporating RGB-D representations of virtual reality objects into the original NYU depth v2 images. We specifically did not match each generated virtual object with an appropriate texture and a suitable location within the real-world image. Instead, an assignment of texture, location, lighting, and other rendering parameters was randomized to maximize a diversity of the training data, and to show that it is randomness that can improve the generalizing ability of a dataset. By conducting extensive experiments with our virtually modified dataset and validating on the original NYU depth v2 and iBims-1 benchmarks, we show that ANYU improves the monocular depth estimation performance and generalization of deep neural networks with considerably different architectures, especially for the current state-of-the-art VPD model. To the best of our knowledge, this is the first work that augments a real-world dataset with randomly generated virtual 3D objects for monocular depth estimation. We make our ANYU dataset publicly available in two training configurations with 10% and 100% additional synthetically enriched RGB-D pairs of training images, respectively, for efficient training and empirical exploration of virtual augmentation at https://github.com/ABrain-One/ANYU

📄 PDF Abstract BibTeX arXiv:2404.09469

Code (1)

abrain-one/anyu 공식 구현

Tasks

Depth EstimationMonocular Depth Estimation

Similar Papers 제목 키워드 기반

MonoMAE: Enhancing Monocular 3D Detection through Depth-Aware Masked Autoencoders

2024-05-13 · Xueying Jiang, Sheng Jin, Xiaoqin Zhang, Ling Shao 외

Monocular 3D object detection aims for precise 3D localization and identification of objects from a single-view image. Despite its recent progress, it often struggles while handling pervasive object occlusions that tend …

3D Object DetectionMonocular 3D Object DetectionObjectobject-detection+1

D4D: An RGBD diffusion model to boost monocular depth estimation

2024-03-12 · L. Papa, P. Russo, I. Amerini

Ground-truth RGBD data are fundamental for a wide range of computer vision applications; however, those labeled samples are difficult to collect and time-consuming to produce. A common solution to overcome this lack of d…

Depth EstimationMonocular Depth Estimation

Fin3R: Fine-tuning Feed-forward 3D Reconstruction Models via Monocular Knowledge Distillation

2025-11-27 · Weining Ren, Hongjun Wang, Xiao Tan, Kai Han arxiv

We present Fin3R, a simple, effective, and general fine-tuning method for feed-forward 3D reconstruction models. The family of feed-forward reconstruction model regresses pointmap of all input images to a reference frame…

Knowledge Distillation3D Reconstruction

Towards Generalization Across Depth for Monocular 3D Object Detection

2019-12-17 · ECCV 2020 8 · Andrea Simonelli, Samuel Rota Bulò, Lorenzo Porzi, Elisa Ricci 외

While expensive LiDAR and stereo camera rigs have enabled the development of successful 3D object detection methods, monocular RGB-only approaches lag much behind. This work advances the state of the art by introducing M…

3D Object DetectionMonocular 3D Object DetectionObjectobject-detection+1

VFMM3D: Releasing the Potential of Image by Vision Foundation Model for Monocular 3D Object Detection

2024-04-15 · Bonan Ding, Jin Xie, Jing Nie, Jiale Cao 외

Due to its cost-effectiveness and widespread availability, monocular 3D object detection, which relies solely on a single camera during inference, holds significant importance across various applications, including auton…

3D Object DetectionAutonomous DrivingMonocular 3D Object DetectionObject+2