paper-with-me

홈 › Papers

Robust Monocular SLAM for Egocentric Videos

2017-07-18 · Suvam Patra, Kartikeya Gupta, Faran Ahmad, Chetan Arora, Subhashis Banerjee

Regardless of the tremendous progress, a truly general purpose pipeline for Simultaneous Localization and Mapping (SLAM) remains a challenge. We investigate the reported failure of state of the art (SOTA) SLAM techniques on egocentric videos. We find that the dominant 3D rotations, low parallax between successive frames, and primarily forward motion in egocentric videos are the most common causes of failures. The incremental nature of SOTA SLAM, in the presence of unreliable pose and 3D estimates in egocentric videos, with no opportunities for global loop closures, generates drifts and leads to the eventual failures of such techniques. Taking inspiration from batch mode Structure from Motion (SFM) techniques, we propose to solve SLAM as an SFM problem over the sliding temporal windows. This makes the problem well constrained. Further, we propose to initialize the camera poses using 2D rotation averaging, followed by translation averaging before structure estimation using bundle adjustment. This helps in stabilizing the camera poses when 3D estimates are not reliable. We show that the proposed SLAM technique, incorporating the two key ideas works successfully for long, shaky egocentric videos where other SOTA techniques have been reported to fail. Qualitative and quantitative comparisons on publicly available egocentric video datasets validate our results.

📄 PDF Abstract BibTeX arXiv:1707.05564

Code (0)

등록된 구현이 없습니다.

Tasks

Simultaneous Localization and Mapping

Similar Papers 제목 키워드 기반

Computing Egomotion with Local Loop Closures for Egocentric Videos

2017-01-17 · Suvam Patra, Himanshu Aggarwal, Himani Arora, Chetan Arora 외

Finding the camera pose is an important step in many egocentric video applications. It has been widely reported that, state of the art SLAM algorithms fail on egocentric videos. In this paper, we propose a robust method …

Camera Pose EstimationDepth EstimationPose Estimation

Pseudo RGB-D for Self-Improving Monocular SLAM and Depth Prediction

2020-04-22 · ECCV 2020 8 · Lokender Tiwari, Pan Ji, Quoc-Huy Tran, Bingbing Zhuang 외

Classical monocular Simultaneous Localization And Mapping (SLAM) and the recently emerging convolutional neural networks (CNNs) for monocular depth prediction represent two largely disjoint approaches towards building a …

Depth EstimationDepth PredictionSimultaneous Localization and Mapping

4D Human Body Capture from Egocentric Video via 3D Scene Grounding

2020-11-26 · Miao Liu, Dexin Yang, Yan Zhang, Zhaopeng Cui 외

We introduce a novel task of reconstructing a time series of second-person 3D human body meshes from monocular egocentric videos. The unique viewpoint and rapid embodied camera motion of egocentric videos raise additiona…

Time SeriesTime Series Analysis

Dyn-HaMR: Recovering 4D Interacting Hand Motion from a Dynamic Camera

2024-12-17 · CVPR 2025 1 · Zhengdi Yu, Stefanos Zafeiriou, Tolga Birdal

We propose Dyn-HaMR, to the best of our knowledge, the first approach to reconstruct 4D global hand motion from monocular videos recorded by dynamic cameras in the wild. Reconstructing accurate 3D hand meshes from monocu…

Simultaneous Localization and Mapping

Self-Supervised Monocular 4D Scene Reconstruction for Egocentric Videos

2024-11-14 · Chengbo Yuan, Geng Chen, Li Yi, Yang Gao

Egocentric videos provide valuable insights into human interactions with the physical world, which has sparked growing interest in the computer vision and robotics communities. A critical challenge in fully understanding…

4D reconstructionSelf-Supervised LearningZero-shot Generalization