paper-with-me

Papers

6D-Diff: A Keypoint Diffusion Framework for 6D Object Pose Estimation

2023-12-29 · CVPR 2024 1 · Li Xu, Haoxuan Qu, Yujun Cai, Jun Liu

Estimating the 6D object pose from a single RGB image often involves noise and indeterminacy due to challenges such as occlusions and cluttered backgrounds. Meanwhile, diffusion models have shown appealing performance in generating high-quality images from random noise with high indeterminacy through step-by-step denoising. Inspired by their denoising capability, we propose a novel diffusion-based framework (6D-Diff) to handle the noise and indeterminacy in object pose estimation for better performance. In our framework, to establish accurate 2D-3D correspondence, we formulate 2D keypoints detection as a reverse diffusion (denoising) process. To facilitate such a denoising process, we design a Mixture-of-Cauchy-based forward diffusion process and condition the reverse process on the object features. Extensive experiments on the LM-O and YCB-V datasets demonstrate the effectiveness of our framework.

📄 PDF Abstract BibTeX arXiv:2401.00029

Code (0)

등록된 구현이 없습니다.

Tasks

6D Pose Estimation using RGBDenoisingObjectPose Estimation

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

KeyPointDiffuser: Unsupervised 3D Keypoint Learning via Latent Diffusion Models

2025-12-03 · Rhys Newbury, Juyan Zhang, Tin Tran, Hanna Kurniawati 외 arxiv

Understanding and representing the structure of 3D objects in an unsupervised manner remains a core challenge in computer vision and graphics. Most existing unsupervised keypoint methods are not designed for unconditiona…

Unsupervised Monocular 3D Keypoint Discovery from Multi-View Diffusion Priors

2025-07-16 · Subin Jeon, In Cho, Junyoung Hong, Seon Joo Kim

This paper introduces KeyDiff3D, a framework for unsupervised monocular 3D keypoints estimation that accurately predicts 3D keypoints from a single image. While previous methods rely on manual annotations or calibrated m…

Unpaired Object-Level SAR-to-Optical Image Translation for Aircraft with Keypoints-Guided Diffusion Models

2025-03-25 · Ruixi You, Hecheng Jia, Feng Xu

Synthetic Aperture Radar (SAR) imagery provides all-weather, all-day, and high-resolution imaging capabilities but its unique imaging mechanism makes interpretation heavily reliant on expert knowledge, limiting interpret…

TranslationZero-shot Generalization

UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation

2025-07-03 · Qin Guo, Ailing Zeng, Dongxu Yue, Ceyuan Yang 외

Although significant advancements have been achieved in the progress of keypoint-guided Text-to-Image diffusion models, existing mainstream keypoint-guided models encounter challenges in controlling the generation of mor…

Image Generation

Foundation Model-Driven Grasping of Unknown Objects via Center of Gravity Estimation

2025-07-25 · Kang Xiangli, Yage He, Xianwu Gong, Zehan Liu 외 arxiv

This study presents a grasping method for objects with uneven mass distribution by leveraging diffusion models to localize the center of gravity (CoG) on unknown objects. In robotic grasping, CoG deviation often leads to…

Robotic Grasping