Pos3R: 6D Pose Estimation for Unseen Objects Made Easy
Foundation models have significantly reduced the need for task-specific training, while also enhancing generalizability. However, state-of-the-art 6D pose estimators either require further training with pose supervision or neglect advances obtainable from 3D foundation models. The latter is a missed opportunity, since these models are better equipped to predict 3D-consistent features, which are of significant utility for the pose estimation task. To address this gap, we propose Pos3R, a method for estimating the 6D pose of any object from a single RGB image, making extensive use of a 3D reconstruction foundation model and requiring no additional training. We identify template selection as a particular bottleneck for existing methods that is significantly alleviated by the use of a 3D model, which can more easily distinguish between template poses than a 2D model. Despite its simplicity, Pos3R achieves competitive performance on the Benchmark for 6D Object Pose Estimation (BOP), matching or surpassing existing refinement-free methods. Additionally, Pos3R integrates seamlessly with render-and-compare refinement techniques, demonstrating adaptability for high-precision applications.
Code (0)
등록된 구현이 없습니다.
Tasks
3D Reconstruction6D Pose Estimation6D Pose Estimation using RGBPose EstimationSimilar Papers 제목 키워드 기반
Unseen Object 6D Pose Estimation: A Benchmark and Baselines
Estimating the 6D pose for unseen objects is in great demand for many real-world applications. However, current state-of-the-art pose estimation methods can only handle objects that are previously trained. In this paper,…
6D Pose EstimationPose EstimationA Large Contextual Dataset for Classification, Detection and Counting of Cars with Deep Learning
We have created a large diverse set of cars from overhead images, which are useful for training a deep learner to binary classify, detect and count them. The dataset and all related material will be made publically avail…
Density EstimationGeneral ClassificationLatentFusion: End-to-End Differentiable Reconstruction and Rendering for Unseen Object Pose Estimation
Current 6D object pose estimation methods usually require a 3D model for each object. These methods also require additional training in order to incorporate new objects. As a result, they are difficult to scale to a larg…
6D Pose Estimation6D Pose Estimation using RGBObjectPose EstimationEasyInsert: A Data-Efficient and Generalizable Insertion Policy
Insertion task is highly challenging that requires robots to operate with exceptional precision in cluttered environments. Existing methods often have poor generalization capabilities. They typically function in restrict…
Pose PredictionZero-shot GeneralizationVisual Imitation Made Easy
Visual imitation learning provides a framework for learning complex manipulation behaviors by leveraging human demonstrations. However, current interfaces for imitation such as kinesthetic teaching or teleoperation prohi…
Imitation Learning