paper-with-me

홈 › Papers

Mono-STAR: Mono-camera Scene-level Tracking and Reconstruction

2023-01-30 · Haonan Chang, Dhruv Metha Ramesh, Shijie Geng, Yuqiu Gan, Abdeslam Boularias

We present Mono-STAR, the first real-time 3D reconstruction system that simultaneously supports semantic fusion, fast motion tracking, non-rigid object deformation, and topological change under a unified framework. The proposed system solves a new optimization problem incorporating optical-flow-based 2D constraints to deal with fast motion and a novel semantic-aware deformation graph (SAD-graph) for handling topology change. We test the proposed system under various challenging scenes and demonstrate that it significantly outperforms existing state-of-the-art methods.

📄 PDF Abstract BibTeX arXiv:2301.13244

Code (1)

changhaonan/mono-star-demo 공식 구현

Tasks

3D ReconstructionOptical Flow Estimation

Methods 이 논문이 사용한 방법론

Test 설명 없음

Similar Papers 제목 키워드 기반

GFlow: Recovering 4D World from Monocular Video

2024-05-28 · Shizun Wang, Xingyi Yang, Qiuhong Shen, Zhenxiang Jiang 외

Recovering 4D world from monocular video is a crucial yet challenging task. Conventional methods usually rely on the assumptions of multi-view videos, known camera parameters, or static scenes. In this paper, we relax al…

4D reconstructionNovel View SynthesisOptical Flow Estimation

State of the Art in Dense Monocular Non-Rigid 3D Reconstruction

2022-10-27 · Edith Tretschk, Navami Kairanda, Mallikarjun B R, Rishabh Dabral 외

3D reconstruction of deformable (or non-rigid) scenes from a set of monocular 2D image observations is a long-standing and actively researched area of computer vision and graphics. It is an ill-posed inverse problem, sin…

3D Reconstruction

Mono-hydra: Real-time 3D scene graph construction from monocular camera input with IMU

2023-08-10 · U. V. B. L. Udugama, G. Vosselman, F. Nex

The ability of robots to autonomously navigate through 3D environments depends on their comprehension of spatial concepts, ranging from low-level geometry to high-level semantics, such as objects, places, and buildings. …

Decision MakingGPUgraph constructionNavigate+1

Self-Supervised Monocular 4D Scene Reconstruction for Egocentric Videos

2024-11-14 · Chengbo Yuan, Geng Chen, Li Yi, Yang Gao

Egocentric videos provide valuable insights into human interactions with the physical world, which has sparked growing interest in the computer vision and robotics communities. A critical challenge in fully understanding…

4D reconstructionSelf-Supervised LearningZero-shot Generalization

Comparative Study of Vision-Based Metric Measurement for Large-Scale Planar Scenes

2026-05-26 · ZhiXin Sun arxiv

Vision-based metric distance and area measurement remains challenging in large-scale outdoor environments due to long-range sensing, camera zoom, and unstable imaging conditions. This work studies planar metric measureme…

Image Stitching