paper-with-me

Papers

MK-Pose: Category-Level Object Pose Estimation via Multimodal-Based Keypoint Learning

2025-07-09 · Yifan Yang, Peili Song, Enfan Lan, Dong Liu, Jingtai Liu

Category-level object pose estimation, which predicts the pose of objects within a known category without prior knowledge of individual instances, is essential in applications like warehouse automation and manufacturing. Existing methods relying on RGB images or point cloud data often struggle with object occlusion and generalization across different instances and categories. This paper proposes a multimodal-based keypoint learning framework (MK-Pose) that integrates RGB images, point clouds, and category-level textual descriptions. The model uses a self-supervised keypoint detection module enhanced with attention-based query generation, soft heatmap matching and graph-based relational modeling. Additionally, a graph-enhanced feature fusion module is designed to integrate local geometric information and global context. MK-Pose is evaluated on CAMERA25 and REAL275 dataset, and is further tested for cross-dataset capability on HouseCat6D dataset. The results demonstrate that MK-Pose outperforms existing state-of-the-art methods in both IoU and average precision without shape priors. Codes will be released at \href{https://github.com/yangyifanYYF/MK-Pose}{https://github.com/yangyifanYYF/MK-Pose}.

📄 PDF Abstract BibTeX arXiv:2507.06662

Code (0)

등록된 구현이 없습니다.

Tasks

Keypoint DetectionPose Estimation

Methods 이 논문이 사용한 방법론

Heatmap 설명 없음

Similar Papers 제목 키워드 기반

iCaps: Iterative Category-level Object Pose and Shape Estimation

2021-12-31 · Xinke Deng, Junyi Geng, Timothy Bretl, Yu Xiang 외

This paper proposes a category-level 6D object pose and shape estimation approach iCaps, which allows tracking 6D poses of unseen objects in a category and estimating their 3D shapes. We develop a category-level auto-enc…

Object

TransNet: Category-Level Transparent Object Pose Estimation

2022-08-22 · Huijie Zhang, Anthony Opipari, Xiaotong Chen, Jiyue Zhu 외

Transparent objects present multiple distinct challenges to visual perception systems. First, their lack of distinguishing visual features makes transparent objects harder to detect and localize than opaque objects. Even…

Depth CompletionObjectPose EstimationSurface Normal Estimation+1

RCGNet: RGB-based Category-Level 6D Object Pose Estimation with Geometric Guidance

2025-08-19 · Sheng Yu, Di-Hua Zhai, Yuanqing Xia arxiv

While most current RGB-D-based category-level object pose estimation methods achieve strong performance, they face significant challenges in scenes lacking depth information. In this paper, we propose a novel category-le…

Pose Estimation

TransNet: Transparent Object Manipulation Through Category-Level Pose Estimation

2023-07-23 · Huijie Zhang, Anthony Opipari, Xiaotong Chen, Jiyue Zhu 외

Transparent objects present multiple distinct challenges to visual perception systems. First, their lack of distinguishing visual features makes transparent objects harder to detect and localize than opaque objects. Even…

Depth CompletionObjectPose EstimationSurface Normal Estimation+1

Category-Level and Open-Set Object Pose Estimation for Robotics

2025-04-28 · Peter Hönig, Matthias Hirschmanner, Markus Vincze

Object pose estimation enables a variety of tasks in computer vision and robotics, including scene understanding and robotic grasping. The complexity of a pose estimation task depends on the unknown variables related to …

6D Pose Estimation6D Pose Estimation using RGBObjectPose Estimation+2