paper-with-me

홈 › Papers

DTF-Net: Category-Level Pose Estimation and Shape Reconstruction via Deformable Template Field

2023-08-04 · Haowen Wang, Zhipeng Fan, Zhen Zhao, Zhengping Che, Zhiyuan Xu, Dong Liu, Feifei Feng, Yakun Huang, XIUQUAN QIAO, Jian Tang

Estimating 6D poses and reconstructing 3D shapes of objects in open-world scenes from RGB-depth image pairs is challenging. Many existing methods rely on learning geometric features that correspond to specific templates while disregarding shape variations and pose differences among objects in the same category. As a result, these methods underperform when handling unseen object instances in complex environments. In contrast, other approaches aim to achieve category-level estimation and reconstruction by leveraging normalized geometric structure priors, but the static prior-based reconstruction struggles with substantial intra-class variations. To solve these problems, we propose the DTF-Net, a novel framework for pose estimation and shape reconstruction based on implicit neural fields of object categories. In DTF-Net, we design a deformable template field to represent the general category-wise shape latent features and intra-category geometric deformation features. The field establishes continuous shape correspondences, deforming the category template into arbitrary observed instances to accomplish shape reconstruction. We introduce a pose regression module that shares the deformation features and template codes from the fields to estimate the accurate 6D pose of each object in the scene. We integrate a multi-modal representation extraction module to extract object features and semantic masks, enabling end-to-end inference. Moreover, during training, we implement a shape-invariant training strategy and a viewpoint sampling method to further enhance the model's capability to extract object pose features. Extensive experiments on the REAL275 and CAMERA25 datasets demonstrate the superiority of DTF-Net in both synthetic and real scenes. Furthermore, we show that DTF-Net effectively supports grasping tasks with a real robot arm.

📄 PDF Abstract BibTeX arXiv:2308.02239

Code (0)

등록된 구현이 없습니다.

Tasks

ObjectPose Estimation

Similar Papers 제목 키워드 기반

Leveraging SE(3) Equivariance for Self-Supervised Category-Level Object Pose Estimation

2021-10-30 · NeurIPS 2021 12 · Xiaolong Li, Yijia Weng, Li Yi, Leonidas Guibas 외

Category-level object pose estimation aims to find 6D object poses of previously unseen object instances from known categories without access to object CAD models. To reduce the huge amount of pose annotations needed for…

ObjectPose EstimationSelf-Supervised Learning

Leveraging SE(3) Equivariance for Self-supervised Category-Level Object Pose Estimation from Point Clouds

2021-05-21 · NeurIPS 2021 12 · Xiaolong Li, Yijia Weng, Li Yi, Leonidas Guibas 외

Category-level object pose estimation aims to find 6D object poses of previously unseen object instances from known categories without access to object CAD models. To reduce the huge amount of pose annotations needed fo…

ObjectPose EstimationSelf-Supervised Learning

Object Level Depth Reconstruction for Category Level 6D Object Pose Estimation From Monocular RGB Image

2022-04-04 · Zhaoxin Fan, Zhenbo Song, Jian Xu, Zhicheng Wang 외

Recently, RGBD-based category-level 6D object pose estimation has achieved promising improvement in performance, however, the requirement of depth information prohibits broader applications. In order to relieve this prob…

6D Pose Estimation using RGBObjectPose Estimation

Category-level Meta-learned NeRF Priors for Efficient Object Mapping

2025-03-03 · Saad Ejaz, Hriday Bavle, Laura Ribeiro, Holger Voos 외

In 3D object mapping, category-level priors enable efficient object reconstruction and canonical pose estimation, requiring only a single prior per semantic category (e.g., chair, book, laptop). Recently, DeepSDF has pre…

GPUMeta-LearningNeRFObject+2

Diffusion-Driven Self-Supervised Learning for Shape Reconstruction and Pose Estimation

2024-03-19 · Jingtao Sun, Yaonan Wang, Mingtao Feng, Chao Ding 외

Fully-supervised category-level pose estimation aims to determine the 6-DoF poses of unseen instances from known categories, requiring expensive mannual labeling costs. Recently, various self-supervised category-level po…

Pose EstimationSelf-Supervised Learning