paper-with-me

홈 › Papers

Real-Time RGB-D Camera Pose Estimation in Novel Scenes using a Relocalisation Cascade

2018-10-29 · Tommaso Cavallari, Stuart Golodetz, Nicholas A. Lord, Julien Valentin, Victor A. Prisacariu, Luigi Di Stefano, Philip H. S. Torr

Camera pose estimation is an important problem in computer vision. Common techniques either match the current image against keyframes with known poses, directly regress the pose, or establish correspondences between keypoints in the image and points in the scene to estimate the pose. In recent years, regression forests have become a popular alternative to establish such correspondences. They achieve accurate results, but have traditionally needed to be trained offline on the target scene, preventing relocalisation in new environments. Recently, we showed how to circumvent this limitation by adapting a pre-trained forest to a new scene on the fly. The adapted forests achieved relocalisation performance that was on par with that of offline forests, and our approach was able to estimate the camera pose in close to real time. In this paper, we present an extension of this work that achieves significantly better relocalisation performance whilst running fully in real time. To achieve this, we make several changes to the original approach: (i) instead of accepting the camera pose hypothesis without question, we make it possible to score the final few hypotheses using a geometric approach and select the most promising; (ii) we chain several instantiations of our relocaliser together in a cascade, allowing us to try faster but less accurate relocalisation first, only falling back to slower, more accurate relocalisation as necessary; and (iii) we tune the parameters of our cascade to achieve effective overall performance. These changes allow us to significantly improve upon the performance our original state-of-the-art method was able to achieve on the well-known 7-Scenes and Stanford 4 Scenes benchmarks. As additional contributions, we present a way of visualising the internal behaviour of our forests and show how to entirely circumvent the need to pre-train a forest on a generic scene.

📄 PDF Abstract BibTeX arXiv:1810.12163

Code (1)

torrvision/spaint 공식 구현

Tasks

Camera Pose EstimationPose Estimation

Similar Papers 제목 키워드 기반

RealMonoDepth: Self-Supervised Monocular Depth Estimation for General Scenes

2020-04-14 · Mertalp Ocal, Armin Mustafa

We present a generalised self-supervised learning approach for monocular estimation of the real depth across scenes with diverse depth ranges from 1--100s of meters. Existing supervised methods for monocular depth estima…

Depth EstimationMonocular Depth EstimationSelf-Supervised Learning

Look Gauss, No Pose: Novel View Synthesis using Gaussian Splatting without Accurate Pose Initialization

2024-10-11 · Christian Schmidt, Jens Piekenbrinck, Bastian Leibe

3D Gaussian Splatting has recently emerged as a powerful tool for fast and accurate novel-view synthesis from a set of posed input images. However, like most novel-view synthesis approaches, it relies on accurate camera …

Camera Pose EstimationNovel View SynthesisPose Estimation

MegaSaM: Accurate, Fast and Robust Structure and Motion from Casual Dynamic Videos

2025-01-01 · CVPR 2025 1 · Zhengqi Li, Richard Tucker, Forrester Cole, Qianqian Wang 외

We present a system that allows for accurate, fast, and robust estimation of camera parameters and depth maps from casual monocular videos of dynamic scenes. Most conventional structure from motion and monocular SLAM…

Depth Estimation

Optical Flow Estimation for Spiking Camera

2021-10-08 · CVPR 2022 1 · Liwen Hu, Rui Zhao, Ziluo Ding, Lei Ma 외

As a bio-inspired sensor with high temporal resolution, the spiking camera has an enormous potential in real applications, especially for motion estimation in high-speed scenes. However, frame-based and event-based metho…

Event-based visionMotion EstimationOptical Flow Estimation

MegaSaM: Accurate, Fast, and Robust Structure and Motion from Casual Dynamic Videos

2024-12-05 · Zhengqi Li, Richard Tucker, Forrester Cole, Qianqian Wang 외

We present a system that allows for accurate, fast, and robust estimation of camera parameters and depth maps from casual monocular videos of dynamic scenes. Most conventional structure from motion and monocular SLAM tec…

Depth Estimation