paper-with-me

홈 › Papers

CRISP: Object Pose and Shape Estimation with Test-Time Adaptation

2024-12-02 · CVPR 2025 1 · Jingnan Shi, Rajat Talak, Harry Zhang, David Jin, Luca Carlone

We consider the problem of estimating object pose and shape from an RGB-D image. Our first contribution is to introduce CRISP, a category-agnostic object pose and shape estimation pipeline. The pipeline implements an encoder-decoder model for shape estimation. It uses FiLM-conditioning for implicit shape reconstruction and a DPT-based network for estimating pose-normalized points for pose estimation. As a second contribution, we propose an optimization-based pose and shape corrector that can correct estimation errors caused by a domain gap. Observing that the shape decoder is well behaved in the convex hull of known shapes, we approximate the shape decoder with an active shape model, and show that this reduces the shape correction problem to a constrained linear least squares problem, which can be solved efficiently by an interior point algorithm. Third, we introduce a self-training pipeline to perform self-supervised domain adaptation of CRISP. The self-training is based on a correct-and-certify approach, which leverages the corrector to generate pseudo-labels at test time, and uses them to self-train CRISP. We demonstrate CRISP (and the self-training) on YCBV, SPE3R, and NOCS datasets. CRISP shows high performance on all the datasets. Moreover, our self-training is capable of bridging a large domain gap. Finally, CRISP also shows an ability to generalize to unseen objects. Code and pre-trained models will be available on https://web.mit.edu/sparklab/research/crisp_object_pose_shape/.

📄 PDF Abstract BibTeX arXiv:2412.01052

Code (0)

등록된 구현이 없습니다.

Tasks

DecoderDomain AdaptationPose EstimationTest-time Adaptation

Similar Papers 제목 키워드 기반

Object Pose and Shape Estimation for Grasping: Does it Work?

2026-05-26 · Pavan Karke, Kushal Shah, Gaurav Singh, Md Faizal Karim 외 arxiv

The problem of object pose and shape estimation has seen key advancements lately. Encoder-decoder (e.g., SAM3D, LRM, CRISP) and diffusion-based models (e.g., InstantMesh, Zero123, SceneComplete) have shown category-agnos…

Hydra++: Real-Time Hierarchical 3D Scene Graph Construction With Object-Level Shape Estimation

2026-07-10 · Hyungtae Lim, Nathan Hughes, Xihang Yu, Ruihan Xu 외 arxiv

3D scene graphs provide a hierarchical abstraction of environments by encoding spatial entities, such as objects and places, and their relationships. However, existing scene graph systems model object geometry coarsely, …

Point Clouds

Deep Crisp Boundaries: From Boundaries to Higher-level Tasks

2018-01-08 · Yupei Wang, Xin Zhao, Yin Li, Kaiqi Huang

Edge detection has made significant progress with the help of deep Convolutional Networks (ConvNet). These ConvNet based edge detectors have approached human level performance on standard benchmarks. We provide a systema…

Edge DetectionObject Proposal GenerationOptical Flow EstimationSemantic Segmentation

Fuzzy c-Shape: A new algorithm for clustering finite time series waveforms

2016-08-03 · Fateme Fahiman, Jame C. Bezdek, Sarah M. Erfani, Christopher Leckie 외

The existence of large volumes of time series data in many applications has motivated data miners to investigate specialized methods for mining time series data. Clustering is a popular data mining method due to its powe…

ClusteringTime SeriesTime Series Analysis

CRISP: A Probabilistic Model for Individual-Level COVID-19 Infection Risk Estimation Based on Contact Data

2020-06-09 · Ralf Herbrich, Rajeev Rastogi, Roland Vollgraf

We present CRISP (COVID-19 Risk Score Prediction), a probabilistic graphical model for COVID-19 infection spread through a population based on the SEIR model where we assume access to (1) mutual contacts between pairs of…

Time Series Analysis