paper-with-me

Papers

MV-ROPE: Multi-view Constraints for Robust Category-level Object Pose and Size Estimation

2023-08-17 · Jiaqi Yang, Yucong Chen, Xiangting Meng, Chenxin Yan, Min Li, Ran Cheng, Lige Liu, Tao Sun, Laurent Kneip

Recently there has been a growing interest in category-level object pose and size estimation, and prevailing methods commonly rely on single view RGB-D images. However, one disadvantage of such methods is that they require accurate depth maps which cannot be produced by consumer-grade sensors. Furthermore, many practical real-world situations involve a moving camera that continuously observes its surroundings, and the temporal information of the input video streams is simply overlooked by single-view methods. We propose a novel solution that makes use of RGB video streams. Our framework consists of three modules: a scale-aware monocular dense SLAM solution, a lightweight object pose predictor, and an object-level pose graph optimizer. The SLAM module utilizes a video stream and additional scale-sensitive readings to estimate camera poses and metric depth. The object pose predictor then generates canonical object representations from RGB images. The object pose is estimated through geometric registration of these canonical object representations with estimated object depth points. All per-view estimates finally undergo optimization within a pose graph, culminating in the output of robust and accurate canonical object poses. Our experimental results demonstrate that when utilizing public dataset sequences with high-quality depth information, the proposed method exhibits comparable performance to state-of-the-art RGB-D methods. We also collect and evaluate on new datasets containing depth maps of varying quality to further quantitatively benchmark the proposed method alongside previous RGB-D based methods. We demonstrate a significant advantage in scenarios where depth input is absent or the quality of depth sensing is limited.

📄 PDF Abstract BibTeX arXiv:2308.08856

Code (0)

등록된 구현이 없습니다.

Tasks

Depth EstimationObject

Similar Papers 제목 키워드 기반

Toward General Object-level Mapping from Sparse Views with 3D Diffusion Priors

2024-10-07 · Ziwei Liao, Binbin Xu, Steven L. Waslander

Object-level mapping builds a 3D map of objects in a scene with detailed shapes and poses from multi-view sensor observations. Conventional methods struggle to build complete shapes and estimate accurate poses due to par…

Object

KineDiff3D: Kinematic-Aware Diffusion for Category-Level Articulated Object Shape Reconstruction and Generation

2025-10-20 · WenBo Xu, Liu Liu, Li Zhang, Ran Zhang 외 arxiv

Articulated objects, such as laptops and drawers, exhibit significant challenges for 3D reconstruction and pose estimation due to their multi-part geometries and variable joint configurations, which introduce structural …

3D ReconstructionPose Estimation

Working Paper: Towards a Category-theoretic Comparative Framework for Artificial General Intelligence

2026-03-30 · Pablo de los Riscos, Fernando J. Corbacho, Michael A. Arbib arxiv

AGI has become the Holly Grail of AI with the promise of level intelligence and the major Tech companies around the world are investing unprecedented amounts of resources in its pursuit. Yet, there does not exist a singl…

Multi-View Hierarchical Graph Neural Network for Sketch-Based 3D Shape Retrieval

2026-04-20 · Hang Cheng, Muyan He, Mingyu Fan, Chengfeng Xie 외 arxiv

Sketch-based 3D shape retrieval (SBSR) aims to retrieve 3D shapes that are consistent with the category of the input hand-drawn sketch. The core challenge of this task lies in two aspects: existing methods typically empl…

Graph Neural Network

VT-3DAD: Cross-Category 3D Anomaly Detection via Visual-Text Normal Space Alignment

2026-06-03 · Zi Wang, Katsuya Hotta, Yawen Zou, Koichiro Kamide 외 arxiv

Few-shot cross-category 3D anomaly detection aims to determine whether an unknown point cloud belongs to a target normal category using only a few normal references. Existing training-based methods usually require catego…

3D Anomaly Detection