paper-with-me

홈 › Papers

SCFlow2: Plug-and-Play Object Pose Refiner with Shape-Constraint Scene Flow

2025-01-01 · CVPR 2025 1 · Qingyuan Wang, Rui Song, Jiaojiao Li, Kerui Cheng, David Ferstl, Yinlin Hu

We introduce SCFlow2, a plug-and-play refinement framework for 6D object pose estimation. Most recent 6D object pose methods rely on refinement to get accurate results. However, most existing refinements either suffer from noises in establishing correspondences, or rely on retraining for novel objects. SCFlow2 is based on the SCFlow model designed for iterative RGB refinement with shape constraint, but formulates the additional depth as a regularization in the iteration via 3D scene flow for RGBD frames. The key design of SCFlow2 is an introduction of geometry constraints into the training of recurrent match network, by combining the rigid-motion embeddings in 3D scene flow and 3D shape prior of the target. We train the refinement network on a combination of dataset Objaverse, GSO and ShapeNet, and demonstrate on BOP datasets with novel objects that, after using our method, the result of most state-of-the-art methods improves significantly, without any retraining or fine-tuning.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

6D Pose Estimation using RGBPose Estimation

Similar Papers 제목 키워드 기반

DetRefiner: Model-Agnostic Detection Refinement with Feature Fusion Transformer

2026-05-11 · Soichiro Okazaki, Tatsuya Sasaki, Hiroki Ohashi arxiv

Open-vocabulary object detection (OVOD) aims to detect both seen and unseen categories, yet existing methods often struggle to generalize to novel objects due to limited integration of global and local contextual cues. W…

Object Detection

Hierarchical Document Refinement for Long-context Retrieval-augmented Generation

2025-05-15 · Jiajie Jin, Xiaoxi Li, Guanting Dong, Yuyao Zhang 외

Real-world RAG applications often encounter long-context input scenarios, where redundant information and noise results in higher inference costs and reduced performance. To address these challenges, we propose LongRefin…

Multi-Task LearningRAGRetrievalRetrieval-augmented Generation

BetterDepth: Plug-and-Play Diffusion Refiner for Zero-Shot Monocular Depth Estimation

2024-07-25 · Xiang Zhang, Bingxin Ke, Hayko Riemenschneider, Nando Metzger 외

By training over large-scale datasets, zero-shot monocular depth estimation (MDE) methods show robust performance in the wild but often suffer from insufficient detail. Although recent diffusion-based MDE approaches exhi…

Depth EstimationMonocular Depth Estimation

TransUPR: A Transformer-based Uncertain Point Refiner for LiDAR Point Cloud Semantic Segmentation

2023-02-16 · Zifan Yu, Meida Chen, Zhikang Zhang, Suya You 외

Common image-based LiDAR point cloud semantic segmentation (LiDAR PCSS) approaches have bottlenecks resulting from the boundary-blurring problem of convolution neural networks (CNNs) and quantitation loss of spherical pr…

Image SegmentationSegmentationSemantic Segmentation

TextRefiner: Internal Visual Feature as Efficient Refiner for Vision-Language Models Prompt Tuning

2024-12-11 · Jingjing Xie, Yuxin Zhang, Jun Peng, Zhaohong Huang 외

Despite the efficiency of prompt learning in transferring vision-language models (VLMs) to downstream tasks, existing methods mainly learn the prompts in a coarse-grained manner where the learned prompt vectors are share…

Prompt Learning