paper-with-me

홈 › Papers

ES-MVSNet: Efficient Framework for End-to-end Self-supervised Multi-View Stereo

2023-08-04 · Qiang Zhou, Chaohui Yu, Jingliang Li, Yuang Liu, Jing Wang, Zhibin Wang

Compared to the multi-stage self-supervised multi-view stereo (MVS) method, the end-to-end (E2E) approach has received more attention due to its concise and efficient training pipeline. Recent E2E self-supervised MVS approaches have integrated third-party models (such as optical flow models, semantic segmentation models, NeRF models, etc.) to provide additional consistency constraints, which grows GPU memory consumption and complicates the model's structure and training pipeline. In this work, we propose an efficient framework for end-to-end self-supervised MVS, dubbed ES-MVSNet. To alleviate the high memory consumption of current E2E self-supervised MVS frameworks, we present a memory-efficient architecture that reduces memory usage by 43% without compromising model performance. Furthermore, with the novel design of asymmetric view selection policy and region-aware depth consistency, we achieve state-of-the-art performance among E2E self-supervised MVS methods, without relying on third-party models for additional consistency signals. Extensive experiments on DTU and Tanks&Temples benchmarks demonstrate that the proposed ES-MVSNet approach achieves state-of-the-art performance among E2E self-supervised MVS methods and competitive performance to many supervised and multi-stage self-supervised methods.

📄 PDF Abstract BibTeX arXiv:2308.02191

Code (0)

등록된 구현이 없습니다.

Tasks

GPUNeRFOptical Flow EstimationSemantic Segmentation

Similar Papers 제목 키워드 기반

RC-MVSNet: Unsupervised Multi-View Stereo with Neural Rendering

2022-03-08 · Di Chang, Aljaž Božič, Tong Zhang, Qingsong Yan 외

Finding accurate correspondences among different views is the Achilles' heel of unsupervised Multi-View Stereo (MVS). Existing methods are built upon the assumption that corresponding pixels share similar photometric fea…

Neural Rendering

Pyramid Multi-view Stereo Net with Self-adaptive View Aggregation

2019-12-06 · ECCV 2020 8 · Hongwei Yi, Zizhuang Wei, Mingyu Ding, Runze Zhang 외

n this paper, we propose an effective and efficient pyramid multi-view stereo (MVS) net with self-adaptive view aggregation for accurate and complete dense point cloud reconstruction. Different from using mean square var…

3D Point Cloud Reconstruction3D ReconstructionDepth EstimationPoint cloud reconstruction

Self-supervised Learning of Depth Inference for Multi-view Stereo

2021-04-07 · CVPR 2021 1 · Jiayu Yang, Jose M. Alvarez, Miaomiao Liu

Recent supervised multi-view depth estimation networks have achieved promising results. Similar to all supervised approaches, these networks require ground-truth data during training. However, collecting a large amount o…

Depth EstimationImage ReconstructionSelf-Supervised Learning

Just a Few Points Are All You Need for Multi-View Stereo: A Novel Semi-Supervised Learning Method for Multi-View Stereo

2021-01-01 · ICCV 2021 10 · Taekyung Kim, Jaehoon Choi, Seokeon Choi, Dongki Jung 외

While learning-based multi-view stereo (MVS) methods have recently shown successful performances in quality and efficiency, limited MVS data hampers generalization to unseen environments. A simple solution is to gene…

3D ReconstructionAll

TransMVSNet: Global Context-aware Multi-view Stereo Network with Transformers

2021-11-29 · CVPR 2022 1 · Yikang Ding, Wentao Yuan, Qingtian Zhu, Haotian Zhang 외

In this paper, we present TransMVSNet, based on our exploration of feature matching in multi-view stereo (MVS). We analogize MVS back to its nature of a feature matching task and therefore propose a powerful Feature Matc…

3D ReconstructionFeature Correlation