paper-with-me

홈 › Papers

Self-supervised Geometric Perception

2021-03-04 · CVPR 2021 1 · Heng Yang, Wei Dong, Luca Carlone, Vladlen Koltun

We present self-supervised geometric perception (SGP), the first general framework to learn a feature descriptor for correspondence matching without any ground-truth geometric model labels (e.g., camera poses, rigid transformations). Our first contribution is to formulate geometric perception as an optimization problem that jointly optimizes the feature descriptor and the geometric models given a large corpus of visual measurements (e.g., images, point clouds). Under this optimization formulation, we show that two important streams of research in vision, namely robust model fitting and deep feature learning, correspond to optimizing one block of the unknown variables while fixing the other block. This analysis naturally leads to our second contribution -- the SGP algorithm that performs alternating minimization to solve the joint optimization. SGP iteratively executes two meta-algorithms: a teacher that performs robust model fitting given learned features to generate geometric pseudo-labels, and a student that performs deep feature learning under noisy supervision of the pseudo-labels. As a third contribution, we apply SGP to two perception problems on large-scale real datasets, namely relative camera pose estimation on MegaDepth and point cloud registration on 3DMatch. We demonstrate that SGP achieves state-of-the-art performance that is on-par or superior to the supervised oracles trained using ground-truth labels.

📄 PDF Abstract BibTeX arXiv:2103.03114

Code (2)

theNded/SGP 공식 구현 pytorch
changliu816/CV-paper-review tf

Tasks

Camera Pose EstimationPoint Cloud RegistrationPose Estimation

Similar Papers 제목 키워드 기반

Unsupervised Joint Learning of Depth, Optical Flow, Ego-motion from Video

2021-05-30 · Jianfeng Li, Junqiao Zhao, Shuangfu Song, Tiantian Feng

Estimating geometric elements such as depth, camera motion, and optical flow from images is an important part of the robot's visual perception. We use a joint self-supervised method to estimate the three geometric elemen…

Depth EstimationOptical Flow EstimationSemantic Segmentation

3D Consistency Optimization for Self-Supervised Monocular Video Depth Estimation

2026-06-14 · Yuanye Liu, Ke Zhang, Junzhe Jiang, Li Zhang 외 arxiv

Reliable monocular video depth estimation is crucial for downstream 3D reasoning and embodied AI in endoscopic navigation. However, existing self-supervised approaches typically treat video frames independently or rely o…

Multi-View 3D ReconstructionDepth Estimation

Learning Point Cloud Geometry as a Statistical Manifold: Theory and Practice

2026-05-11 · Jinwoo Lee, Jiwoo Kim, Woojae Shin, Giseop Kim 외 arxiv

Point clouds are a fundamental representation for robotic perception tasks such as localization, mapping, and object pose estimation. However, LiDAR-acquired point clouds are inherently sparse and non-uniform, providing …

Pose EstimationPoint Clouds

AdaCropFollow: Self-Supervised Online Adaptation for Visual Under-Canopy Navigation

2024-10-16 · Arun N. Sivakumar, Federico Magistri, Mateus V. Gasparino, Jens Behley 외

Under-canopy agricultural robots can enable various applications like precise monitoring, spraying, weeding, and plant manipulation tasks throughout the growing season. Autonomous navigation under the canopy is challengi…

Autonomous Navigation

Learning Optical Flow, Depth, and Scene Flow without Real-World Labels

2022-03-28 · Vitor Guizilini, Kuan-Hui Lee, Rares Ambrus, Adrien Gaidon

Self-supervised monocular depth estimation enables robots to learn 3D perception from raw video streams. This scalable approach leverages projective geometry and ego-motion to learn via view synthesis, assuming the world…

Autonomous DrivingDepth EstimationMonocular Depth EstimationMulti-Task Learning+2