paper-with-me

홈 › Papers

Zero123-6D: Zero-shot Novel View Synthesis for RGB Category-level 6D Pose Estimation

2024-03-21 · Francesco Di Felice, Alberto Remus, Stefano Gasperini, Benjamin Busam, Lionel Ott, Federico Tombari, Roland Siegwart, Carlo Alberto Avizzano

Estimating the pose of objects through vision is essential to make robotic platforms interact with the environment. Yet, it presents many challenges, often related to the lack of flexibility and generalizability of state-of-the-art solutions. Diffusion models are a cutting-edge neural architecture transforming 2D and 3D computer vision, outlining remarkable performances in zero-shot novel-view synthesis. Such a use case is particularly intriguing for reconstructing 3D objects. However, localizing objects in unstructured environments is rather unexplored. To this end, this work presents Zero123-6D, the first work to demonstrate the utility of Diffusion Model-based novel-view-synthesizers in enhancing RGB 6D pose estimation at category-level, by integrating them with feature extraction techniques. Novel View Synthesis allows to obtain a coarse pose that is refined through an online optimization method introduced in this work to deal with intra-category geometric differences. In such a way, the outlined method shows reduction in data requirements, removal of the necessity of depth information in zero-shot category-level 6D pose estimation task, and increased performance, quantitatively demonstrated through experiments on the CO3D dataset.

📄 PDF Abstract BibTeX arXiv:2403.14279

Code (0)

등록된 구현이 없습니다.

Tasks

6D Pose EstimationNovel View SynthesisPose Estimation

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Visual Semantic Segmentation Based on Few/Zero-Shot Learning: An Overview

2022-11-13 · Wenqi Ren, Yang Tang, Qiyu Sun, Chaoqiang Zhao 외

Visual semantic segmentation aims at separating a visual sample into diverse blocks with specific semantic attributes and identifying the category for each block, and it plays a crucial role in environmental perception. …

SegmentationSemantic SegmentationVideo Object SegmentationVideo Semantic Segmentation+1

Multi-View Hierarchical Graph Neural Network for Sketch-Based 3D Shape Retrieval

2026-04-20 · Hang Cheng, Muyan He, Mingyu Fan, Chengfeng Xie 외 arxiv

Sketch-based 3D shape retrieval (SBSR) aims to retrieve 3D shapes that are consistent with the category of the input hand-drawn sketch. The core challenge of this task lies in two aspects: existing methods typically empl…

Graph Neural Network

ZeroNVS: Zero-Shot 360-Degree View Synthesis from a Single Image

2023-10-27 · CVPR 2024 1 · Kyle Sargent, Zizhang Li, Tanmay Shah, Charles Herrmann 외

We introduce a 3D-aware diffusion model, ZeroNVS, for single-image novel view synthesis for in-the-wild scenes. While existing methods are designed for single objects with masked backgrounds, we propose new techniques to…

DiversityNeRFNovel View Synthesis

Zero-1-to-3: Zero-shot One Image to 3D Object

2023-03-20 · ICCV 2023 1 · Ruoshi Liu, Rundi Wu, Basile Van Hoorick, Pavel Tokmakov 외

We introduce Zero-1-to-3, a framework for changing the camera viewpoint of an object given just a single RGB image. To perform novel view synthesis in this under-constrained setting, we capitalize on the geometric priors…

3D ReconstructionImage to 3DNovel View SynthesisSingle-View 3D Reconstruction+1

ACNet: Approaching-and-Centralizing Network for Zero-Shot Sketch-Based Image Retrieval

2021-11-24 · Hao Ren, Ziqiang Zheng, Yang Wu, Hong Lu 외

The huge domain gap between sketches and photos and the highly abstract sketch representations pose challenges for sketch-based image retrieval (\underline{SBIR}). The zero-shot sketch-based image retrieval (\underline{Z…

Image RetrievalRetrievalSketch-Based Image Retrieval