paper-with-me

Papers

MonoInstance: Enhancing Monocular Priors via Multi-view Instance Alignment for Neural Rendering and Reconstruction

2025-03-24 · CVPR 2025 1 · Wenyuan Zhang, Yixiao Yang, Han Huang, Liang Han, Kanle Shi, Yu-Shen Liu, Zhizhong Han

Monocular depth priors have been widely adopted by neural rendering in multi-view based tasks such as 3D reconstruction and novel view synthesis. However, due to the inconsistent prediction on each view, how to more effectively leverage monocular cues in a multi-view context remains a challenge. Current methods treat the entire estimated depth map indiscriminately, and use it as ground truth supervision, while ignoring the inherent inaccuracy and cross-view inconsistency in monocular priors. To resolve these issues, we propose MonoInstance, a general approach that explores the uncertainty of monocular depths to provide enhanced geometric priors for neural rendering and reconstruction. Our key insight lies in aligning each segmented instance depths from multiple views within a common 3D space, thereby casting the uncertainty estimation of monocular depths into a density measure within noisy point clouds. For high-uncertainty areas where depth priors are unreliable, we further introduce a constraint term that encourages the projected instances to align with corresponding instance masks on nearby views. MonoInstance is a versatile strategy which can be seamlessly integrated into various multi-view neural rendering frameworks. Our experimental results demonstrate that MonoInstance significantly improves the performance in both reconstruction and novel view synthesis under various benchmarks.

📄 PDF Abstract BibTeX arXiv:2503.18363

Code (0)

등록된 구현이 없습니다.

Tasks

3D ReconstructionNeural RenderingNovel View Synthesis

Methods 이 논문이 사용한 방법론

ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

NC-SDF: Enhancing Indoor Scene Reconstruction Using Neural SDFs with View-Dependent Normal Compensation

2024-05-01 · CVPR 2024 1 · Ziyi Chen, Xiaolong Wu, Yu Zhang

State-of-the-art neural implicit surface representations have achieved impressive results in indoor scene reconstruction by incorporating monocular geometric priors as additional supervision. However, we have observed th…

3D ReconstructionIndoor Scene Reconstruction

DebSDF: Delving into the Details and Bias of Neural Indoor Scene Reconstruction

2023-08-29 · Yuting Xiao, Jingwei Xu, Zehao Yu, Shenghua Gao

In recent years, the neural implicit surface has emerged as a powerful representation for multi-view surface reconstruction due to its simplicity and state-of-the-art performance. However, reconstructing smooth and detai…

Indoor Scene ReconstructionSurface Reconstruction

Vivid4D: Improving 4D Reconstruction from Monocular Video by Video Inpainting

2025-04-15 · Jiaxin Huang, Sheng Miao, Bangbang Yang, Yuewen Ma 외

Reconstructing 4D dynamic scenes from casually captured monocular videos is valuable but highly challenging, as each timestamp is observed from a single viewpoint. We introduce Vivid4D, a novel approach that enhances 4D …

4D reconstructionVideo Inpainting

MP-SfM: Monocular Surface Priors for Robust Structure-from-Motion

2025-04-28 · CVPR 2025 1 · Zador Pataki, Paul-Edouard Sarlin, Johannes L. Schönberger, Marc Pollefeys

While Structure-from-Motion (SfM) has seen much progress over the years, state-of-the-art systems are prone to failure when facing extreme viewpoint changes in low-overlap, low-parallax or high-symmetry scenarios. Becaus…

MonoMVSNet: Monocular Priors Guided Multi-View Stereo Network

2025-07-15 · Jianfei Jiang, Qiankun Liu, Haochen Yu, Hongyuan Liu 외

Learning-based Multi-View Stereo (MVS) methods aim to predict depth maps for a sequence of calibrated images to recover dense point clouds. However, existing MVS methods often struggle with challenging regions, such as t…

Depth EstimationDepth PredictionMonocular Depth Estimation