paper-with-me

Papers

Structure-from-Motion using Dense CNN Features with Keypoint Relocalization

2018-05-10 · Aji Resindra Widya, Akihiko Torii, Masatoshi Okutomi

Structure from Motion (SfM) using imagery that involves extreme appearance changes is yet a challenging task due to a loss of feature repeatability. Using feature correspondences obtained by matching densely extracted convolutional neural network (CNN) features significantly improves the SfM reconstruction capability. However, the reconstruction accuracy is limited by the spatial resolution of the extracted CNN features which is not even pixel-level accuracy in the existing approach. Providing dense feature matches with precise keypoint positions is not trivial because of memory limitation and computational burden of dense features. To achieve accurate SfM reconstruction with highly repeatable dense features, we propose an SfM pipeline that uses dense CNN features with relocalization of keypoint position that can efficiently and accurately provide pixel-level feature correspondences. Then, we demonstrate on the Aachen Day-Night dataset that the proposed SfM using dense CNN features with the keypoint relocalization outperforms a state-of-the-art SfM (COLMAP using RootSIFT) by a large margin.

📄 PDF Abstract BibTeX arXiv:1805.03879

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

ALIKED: A Lighter Keypoint and Descriptor Extraction Network via Deformable Transformation

2023-04-07 · Xiaoming Zhao, Xingming Wu, Weihai Chen, Peter C. Y. Chen 외

Image keypoints and descriptors play a crucial role in many visual measurement tasks. In recent years, deep neural networks have been widely used to improve the performance of keypoint and descriptor extraction. However,…

3D ReconstructionHomography Estimation

Sparse2Dense: A Keypoint-driven Generative Framework for Human Video Compression and Vertex Prediction

2025-09-27 · Bolin Chen, Ru-Ling Liao, Yan Ye, Jie Chen 외 arxiv

For bandwidth-constrained multimedia applications, simultaneously achieving ultra-low bitrate human video compression and accurate vertex prediction remains a critical challenge, as it demands the harmonization of dynami…

Multi-Task LearningVideo Generation

Pixel-Perfect Structure-from-Motion with Featuremetric Refinement

2021-08-18 · ICCV 2021 10 · Philipp Lindenberger, Paul-Edouard Sarlin, Viktor Larsson, Marc Pollefeys

Finding local features that are repeatable across multiple views is a cornerstone of sparse 3D reconstruction. The classical image matching paradigm detects keypoints per-image once and for all, which can yield poorly-lo…

3D Reconstruction

Hi^2-GSLoc: Dual-Hierarchical Gaussian-Specific Visual Relocalization for Remote Sensing

2025-07-21 · Boni Hu, Zhenyu Xia, Lin Chen, Pengcheng Han 외 arxiv

Visual relocalization, which estimates the 6-degree-of-freedom (6-DoF) camera pose from query images, is fundamental to remote sensing and UAV applications. Existing methods face inherent trade-offs: image-based retrieva…

Computational EfficiencyPose Estimation

TransCamP: Graph Transformer for 6-DoF Camera Pose Estimation

2021-05-28 · Xinyi Li, Haibin Ling

Camera pose estimation or camera relocalization is the centerpiece in numerous computer vision tasks such as visual odometry, structure from motion (SfM) and SLAM. In this paper we propose a neural network approach with …

Camera Pose EstimationCamera RelocalizationComputational EfficiencyPose Estimation+1