paper-with-me

홈 › Papers

GenFlow: Generalizable Recurrent Flow for 6D Pose Refinement of Novel Objects

2024-03-18 · CVPR 2024 1 · Sungphill Moon, Hyeontae Son, Dongcheol Hur, SangWook Kim

Despite the progress of learning-based methods for 6D object pose estimation, the trade-off between accuracy and scalability for novel objects still exists. Specifically, previous methods for novel objects do not make good use of the target object's 3D shape information since they focus on generalization by processing the shape indirectly, making them less effective. We present GenFlow, an approach that enables both accuracy and generalization to novel objects with the guidance of the target object's shape. Our method predicts optical flow between the rendered image and the observed image and refines the 6D pose iteratively. It boosts the performance by a constraint of the 3D shape and the generalizable geometric knowledge learned from an end-to-end differentiable system. We further improve our model by designing a cascade network architecture to exploit the multi-scale correlations and coarse-to-fine refinement. GenFlow ranked first on the unseen object pose estimation benchmarks in both the RGB and RGB-D cases. It also achieves performance competitive with existing state-of-the-art methods for the seen object pose estimation without any fine-tuning.

📄 PDF Abstract BibTeX arXiv:2403.11510

Code (0)

등록된 구현이 없습니다.

Tasks

6D Pose Estimation using RGBObjectOptical Flow EstimationPose Estimation

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

GenFlowRL: Shaping Rewards with Generative Object-Centric Flow in Visual Reinforcement Learning

2025-08-14 · Kelin Yu, Sheng Zhang, Harshit Soora, Furong Huang 외 arxiv

Recent advances have shown that video generation models can enhance robot learning by deriving effective robot actions through inverse dynamics. However, these methods heavily depend on the quality of generated data and …

Reinforcement LearningVideo Generation

GenFlow: Interactive Modular System for Image Generation

2025-06-26 · Duc-Hung Nguyen, Huu-Phuc Huynh, Minh-Triet Tran, Trung-Nghia Le

Generative art unlocks boundless creative possibilities, yet its full potential remains untapped due to the technical expertise required for advanced architectural concepts and computational workflows. To bridge this gap…

Image Generation

Genflow Ad Studio: A Compound AI Architecture for Brand-Aligned, Self-Correcting Video Generation

2026-05-16 · Debanshu Das, Lavi Nigam, Sunil Kumar Jang Bahadur, Gopala Dhar arxiv

Recent advancements in generative video models demonstrate high visual fidelity, yet their integration into enterprise environments is restricted by temporal inconsistencies and severe brand misalignment. Current monolit…

Video Generation

Generalized Gradient Flows with Provable Fixed-Time Convergence and Fast Evasion of Non-Degenerate Saddle Points

2022-12-07 · Mayank Baranwal, Param Budhraja, Vishal Raj, Ashish R. Hota

Gradient-based first-order convex optimization algorithms find widespread applicability in a variety of domains, including machine learning tasks. Motivated by the recent advances in fixed-time stability theory of contin…

SCFlow2: Plug-and-Play Object Pose Refiner with Shape-Constraint Scene Flow

2025-01-01 · CVPR 2025 1 · Qingyuan Wang, Rui Song, Jiaojiao Li, Kerui Cheng 외

We introduce SCFlow2, a plug-and-play refinement framework for 6D object pose estimation. Most recent 6D object pose methods rely on refinement to get accurate results. However, most existing refinements either suffe…

6D Pose Estimation using RGBPose Estimation