paper-with-me

Papers

Every9D-21M: Large-Scale Real-World 9D Canonicalization of Everyday Objects

2026-05-27 · Leonhard Sommer, Emil Akopyan, Adam Kortylewski arxiv

Estimating the 9D pose of everyday objects from a single real-world image remains challenging. This is largely due to the lack of large-scale supervision. Most existing datasets either rely heavily on synthetic renderings or provide limited coverage of real-world objects: the largest real-world 9D pose dataset to date contains only 17K annotated objects across 9 categories. We address this gap with Every9D-21M, a dataset of 9D pose annotations for 21.8M real-world images from 109K object- centric videos spanning 700 everyday object categories - two orders of magnitude larger than prior real-world 9D pose benchmarks in both image and category count. To achieve this scale, we leverage object-centric videos by reconstructing object- level point clouds via multi-view geometry and aligning similar instances into a shared canonical coordinate frame. Canonical poses are manually annotated for only a small set of reference objects (fewer than 0.01% of all images) and propagated to the remaining instances via cross-instance alignment. All propagated canonical poses are then verified from multiple viewpoints. We further introduce cross-category orientation rules that induce category-level symmetries, enabling symmetry-aware evaluation. Beyond establishing dedicated training and evaluation splits as a benchmark for 9D pose foundation models, we show that training on Every9D-21M improves performance on ImageNet3D and PASCAL3D+, and generalizes to HANDAL substantially better than training on ImageNet3D. Data and code are available at https://github.com/GenIntel/Every9D.

📄 PDF Abstract BibTeX arXiv:2605.28270

Code (0)

등록된 구현이 없습니다.

Tasks

Point Clouds

Similar Papers 제목 키워드 기반

Robust Canonicalization through Bootstrapped Data Re-Alignment

2025-10-09 · Johann Schmidt, Sebastian Stober arxiv

Fine-grained visual classification (FGVC) tasks, such as insect and bird identification, demand sensitivity to subtle visual cues while remaining robust to spatial transformations. A key challenge is handling geometric b…

Data Augmentation

DRACO: Weakly Supervised Dense Reconstruction And Canonicalization of Objects

2020-11-25 · Rahul Sajnani, AadilMehdi Sanchawala, Krishna Murthy Jatavallabhula, Srinath Sridhar 외

We present DRACO, a method for Dense Reconstruction And Canonicalization of Object shape from one or more RGB images. Canonical shape reconstruction, estimating 3D object shape in a coordinate space canonicalized for sca…

ObjectPose EstimationTranslation

One-shot 3D Object Canonicalization based on Geometric and Semantic Consistency

2025-01-01 · CVPR 2025 1 · Li Jin, Yujie Wang, Wenzheng Chen, Qiyu Dai 외

3D object canonicalization is a fundamental task, essential for various downstream tasks. Existing methods rely on either cumbersome manual processes or priors learned from extensive, per-category training samples. R…

Object

CanoVerse: 3D Object Scalable Canonicalization and Dataset for Generation and Pose

2026-03-07 · Li Jin, Yuchen Yang, Weikai Chen, Yujie Wang 외 arxiv

3D learning systems implicitly assume that objects occupy a coherent reference frame. Nonetheless, in practice, every asset arrives with an arbitrary global rotation, and models are left to resolve directional ambiguity …

3D Generation

Weisfeiler-Leman Is Incomplete on Simple Spectrum Graphs, so Canonicalize Them

2026-05-22 · Snir Hordan, Nadav Dym, Tim Seppelt arxiv

Graphs with a simple spectrum admit cubic-time isomorphism testing, yet we prove that for every natural number $k$, the $k$-Weisfeiler-Leman ($k$-WL) test cannot distinguish all non-isomorphic graphs with a simple spectr…

Graph Regression