paper-with-me

Papers

$SE(3)$ Equivariant Ray Embeddings for Implicit Multi-View Depth Estimation

2024-11-11 · Yinshuang Xu, Dian Chen, Katherine Liu, Sergey Zakharov, Rares Ambrus, Kostas Daniilidis, Vitor Guizilini

Incorporating inductive bias by embedding geometric entities (such as rays) as input has proven successful in multi-view learning. However, the methods adopting this technique typically lack equivariance, which is crucial for effective 3D learning. Equivariance serves as a valuable inductive prior, aiding in the generation of robust multi-view features for 3D scene understanding. In this paper, we explore the application of equivariant multi-view learning to depth estimation, not only recognizing its significance for computer vision and robotics but also addressing the limitations of previous research. Most prior studies have either overlooked equivariance in this setting or achieved only approximate equivariance through data augmentation, which often leads to inconsistencies across different reference frames. To address this issue, we propose to embed $SE(3)$ equivariance into the Perceiver IO architecture. We employ Spherical Harmonics for positional encoding to ensure 3D rotation equivariance, and develop a specialized equivariant encoder and decoder within the Perceiver IO architecture. To validate our model, we applied it to the task of stereo depth estimation, achieving state of the art results on real-world datasets without explicit geometric constraints or extensive data augmentation.

📄 PDF Abstract BibTeX arXiv:2411.07326

Code (0)

등록된 구현이 없습니다.

Tasks

Data AugmentationDecoderDepth EstimationInductive BiasMULTI-VIEW LEARNINGScene UnderstandingStereo Depth Estimation

Methods 이 논문이 사용한 방법론

Perceiver IO 설명 없음

Similar Papers 제목 키워드 기반

3D Equivariant Graph Implicit Functions

2022-03-31 · Yunlu Chen, Basura Fernando, Hakan Bilen, Matthias Nießner 외

In recent years, neural implicit representations have made remarkable progress in modeling of 3D shapes with arbitrary topology. In this work, we address two key limitations of such representations, in failing to capture…

AeDet: Azimuth-invariant Multi-view 3D Object Detection

2022-11-22 · CVPR 2023 1 · Chengjian Feng, Zequn Jie, Yujie Zhong, Xiangxiang Chu 외

Recent LSS-based multi-view 3D object detection has made tremendous progress, by processing the features in Brid-Eye-View (BEV) via the convolutional detector. However, the typical convolution ignores the radial symmetry…

3D Object DetectionDepth EstimationDepth PredictionObject+2

Occlusion-Invariant Rotation-Equivariant Semi-Supervised Depth Based Cross-View Gait Pose Estimation

2021-09-03 · Xiao Gu, Jianxin Yang, Hanxiao Zhang, Jianing Qiu 외

Accurate estimation of three-dimensional human skeletons from depth images can provide important metrics for healthcare applications, especially for biomechanical gait analysis. However, there exist inherent problems ass…

Pose Estimation

Cross-Domain 3D Equivariant Image Embeddings

2018-12-06 · Carlos Esteves, Avneesh Sud, Zhengyi Luo, Kostas Daniilidis 외

Spherical convolutional networks have been introduced recently as tools to learn powerful feature representations of 3D shapes. Spherical CNNs are equivariant to 3D rotations making them ideally suited to applications wh…

3D Shape ClassificationNovel View SynthesisObjectPose Estimation

Learning Neural Implicit through Volume Rendering with Attentive Depth Fusion Priors

2023-10-17 · NeurIPS 2023 11 · Pengchong Hu, Zhizhong Han

Learning neural implicit representations has achieved remarkable performance in 3D reconstruction from multi-view images. Current methods use volume rendering to render implicit representations into either RGB or depth i…

3D ReconstructionSimultaneous Localization and Mapping