paper-with-me

홈 › Papers

DKPMV: Dense Keypoints Fusion from Multi-View RGB Frames for 6D Pose Estimation of Textureless Objects

2025-10-13 · Jiahong Chen, Jinghao Wang, Zi Wang, Ziwen Wang, Banglei Guan, Qifeng Yu arxiv

6D pose estimation of textureless objects is valuable for industrial robotic applications, yet remains challenging due to the frequent loss of depth information. Current multi-view methods either rely on depth data or insufficiently exploit multi-view geometric cues, limiting their performance. In this paper, we propose DKPMV, a pipeline that achieves dense keypoint-level fusion using only multi-view RGB images as input. We design a three-stage progressive pose optimization strategy that leverages dense multi-view keypoint geometry information. To enable effective dense keypoint fusion, we enhance the keypoint network with attentional aggregation and symmetry-aware training, improving prediction accuracy and resolving ambiguities on symmetric objects. Extensive experiments on the ROBI dataset demonstrate that DKPMV outperforms state-of-the-art multi-view RGB approaches and even surpasses the RGB-D methods in the majority of cases. The code will be available soon.

📄 PDF Abstract BibTeX arXiv:2510.10933

Code (0)

등록된 구현이 없습니다.

Tasks

6D Pose Estimation

Similar Papers 제목 키워드 기반

DiffProxy: Multi-View Human Mesh Recovery via Diffusion-Generated Dense Proxies

2026-01-05 · Renke Wang, Zhenyu Zhang, Ying Tai, Jun Li 외 arxiv

Precise human mesh recovery (HMR) from multi-view images remains challenging: end-to-end methods produce entangled errors hard to localize, while fitting-based methods rely on sparse keypoints that provide limited surfac…

Human Mesh Recovery

Sparse4D: Multi-view 3D Object Detection with Sparse Spatial-Temporal Fusion

2022-11-19 · Xuewu Lin, Tianwei Lin, Zixiang Pei, Lichao Huang 외

Bird-eye-view (BEV) based methods have made great progress recently in multi-view 3D detection task. Comparing with BEV based methods, sparse based methods lag behind in performance, but still have lots of non-negligible…

3D Object Detectionobject-detectionObject DetectionRobust Camera Only 3D Object Detection

Unsupervised Monocular 3D Keypoint Discovery from Multi-View Diffusion Priors

2025-07-16 · Subin Jeon, In Cho, Junyoung Hong, Seon Joo Kim

This paper introduces KeyDiff3D, a framework for unsupervised monocular 3D keypoints estimation that accurately predicts 3D keypoints from a single image. While previous methods rely on manual annotations or calibrated m…

Dense Keypoints via Multiview Supervision

2021-12-01 · NeurIPS 2021 12 · Zhixuan Yu, Haozheng Yu, Long Sha, Sujoy Ganguly 외

This paper presents a new end-to-end semi-supervised framework to learn a dense keypoint detector using unlabeled multiview images. A key challenge lies in finding the exact correspondences between the dense keypoints in …

3D ReconstructionKeypoint Detection

Semi-supervised Dense Keypoints Using Unlabeled Multiview Images

2021-09-20 · Zhixuan Yu, Haozheng Yu, Long Sha, Sujoy Ganguly 외

This paper presents a new end-to-end semi-supervised framework to learn a dense keypoint detector using unlabeled multiview images. A key challenge lies in finding the exact correspondences between the dense keypoints in…

3D ReconstructionKeypoint Detection