paper-with-me

홈 › Papers

Object-Centric Multi-View Aggregation

2020-07-20 · Shubham Tulsiani, Or Litany, Charles R. Qi, He Wang, Leonidas J. Guibas

We present an approach for aggregating a sparse set of views of an object in order to compute a semi-implicit 3D representation in the form of a volumetric feature grid. Key to our approach is an object-centric canonical 3D coordinate system into which views can be lifted, without explicit camera pose estimation, and then combined -- in a manner that can accommodate a variable number of views and is view order independent. We show that computing a symmetry-aware mapping from pixels to the canonical coordinate system allows us to better propagate information to unseen regions, as well as to robustly overcome pose ambiguities during inference. Our aggregate representation enables us to perform 3D inference tasks like volumetric reconstruction and novel view synthesis, and we use these tasks to demonstrate the benefits of our aggregation approach as compared to implicit or camera-centric alternatives.

📄 PDF Abstract BibTeX arXiv:2007.10300

Code (0)

등록된 구현이 없습니다.

Tasks

Camera Pose EstimationNovel View SynthesisObjectPose Estimation

Similar Papers 제목 키워드 기반

EgoSplat: Open-Vocabulary Egocentric Scene Understanding with Language Embedded 3D Gaussian Splatting

2025-03-14 · Di Li, Jie Feng, Jiahao Chen, Weisheng Dong 외

Egocentric scenes exhibit frequent occlusions, varied viewpoints, and dynamic interactions compared to typical scene understanding tasks. Occlusions and varied viewpoints can lead to multi-view semantic inconsistencies, …

Scene UnderstandingSegmentation

Egocentric Audio-Visual Object Localization

2023-03-23 · CVPR 2023 1 · Chao Huang, Yapeng Tian, Anurag Kumar, Chenliang Xu

Humans naturally perceive surrounding scenes by unifying sound and sight in a first-person view. Likewise, machines are advanced to approach human intelligence by learning with multisensory inputs from an egocentric pers…

ObjectObject Localization

ViewFormer: Exploring Spatiotemporal Modeling for Multi-View 3D Occupancy Perception via View-Guided Transformers

2024-05-07 · Jinke Li, Xiao He, Chonghua Zhou, Xiaoqiang Cheng 외

3D occupancy, an advanced perception technology for driving scenarios, represents the entire scene without distinguishing between foreground and background by quantifying the physical space into a grid map. The widely ad…

3D Object Detectionobject-detectionObject Detection

Learning Global Spatial Information for Multi-View Object-Centric Models

2021-09-29 · Yuya Kobayashi, Masahiro Suzuki, Yutaka Matsuo

Recently, several studies have been working on multi-view object-centric models, which predict unobserved views of a scene and infer object-centric representations from several observation views. In general, multi-object…

Novel View SynthesisObject

Learning Object-Centric Representations of Multi-Object Scenes from Multiple Views

2021-11-13 · NeurIPS 2020 12 · Li Nanbo, Cian Eastwood, Robert B. Fisher

Learning object-centric representations of multi-object scenes is a promising approach towards machine intelligence, facilitating high-level reasoning and control from visual sensory data. However, current approaches for…

ObjectScene Understanding