paper-with-me

홈 › Papers

ActMVS: Active Scene Reconstruction with Monocular Multi-View Stereo

2026-05-31 · Guo Pu, Yixuan Han, Zhouhui Lian arxiv

Active scene reconstruction enables robots/UAVs to autonomously plan trajectories and reconstruct environments without costly manual data acquisition. Unlike passive methods, active reconstruction requires real-time construction of high-confidence occupancy maps for collision-free navigation. Existing approaches rely on depth sensors for occupancy map updates, increasing platform cost and weight. To advance spatial intelligence, we aim for a vision-only monocular solution. However, current monocular scene reconstruction methods operate offline and fail to deliver globally consistent dense depth at the frame rates required for robots/UAVs navigation. To bridge this gap, we introduce ActMVS, the first framework for monocular active reconstruction. Our framework integrates a view factor graph construction for informed Multi-View Stereo depth prediction, along with a global depth optimization, to enable the online generation of high-quality, globally consistent dense depth maps. This enables monocular robots/UAVs to maintain reliable occupancy maps for safe trajectory planning during reconstruction. Experiments on Replica datasets demonstrate performance competitive with RGB-D methods. Our code and data are available at https://github.com/TrickyGo/ActMVS.

📄 PDF Abstract BibTeX arXiv:2606.01367

Code (0)

등록된 구현이 없습니다.

Tasks

Trajectory Planning

Similar Papers 제목 키워드 기반

CAT4D: Create Anything in 4D with Multi-View Video Diffusion Models

2024-11-27 · CVPR 2025 1 · Rundi Wu, Ruiqi Gao, Ben Poole, Alex Trevithick 외

We present CAT4D, a method for creating 4D (dynamic 3D) scenes from monocular video. CAT4D leverages a multi-view video diffusion model trained on a diverse combination of datasets to enable novel view synthesis at any s…

4D reconstructionNovel View SynthesisScene Generation

TransformerFusion: Monocular RGB Scene Reconstruction using Transformers

2021-07-05 · NeurIPS 2021 12 · Aljaž Božič, Pablo Palafox, Justus Thies, Angela Dai 외

We introduce TransformerFusion, a transformer-based 3D scene reconstruction approach. From an input monocular RGB video, the video frames are processed by a transformer network that fuses the observations into a volumetr…

3D Reconstruction3D Scene ReconstructionDepth EstimationStereo Depth Estimation+1

State of the Art in Dense Monocular Non-Rigid 3D Reconstruction

2022-10-27 · Edith Tretschk, Navami Kairanda, Mallikarjun B R, Rishabh Dabral 외

3D reconstruction of deformable (or non-rigid) scenes from a set of monocular 2D image observations is a long-standing and actively researched area of computer vision and graphics. It is an ill-posed inverse problem, sin…

3D Reconstruction

UniCon3R: Unified Contact-aware 4D Human-Scene Reconstruction from Monocular Video

2026-04-21 · Tanuj Sur, Shashank Tripathi, Nikos Athanasiou, Ha Linh Nguyen 외 arxiv

We introduce UniCon3R, a unified feed-forward framework for online human-scene 4D reconstruction from monocular video. Current feed-forward human-scene reconstruction methods suffer from artifacts, where bodies float abo…

Fast Monocular Scene Reconstruction with Global-Sparse Local-Dense Grids

2023-05-22 · CVPR 2023 1 · Wei Dong, Chris Choy, Charles Loop, Or Litany 외

Indoor scene reconstruction from monocular images has long been sought after by augmented reality and robotics developers. Recent advances in neural field representations and monocular priors have led to remarkable resul…

Indoor Scene Reconstruction