paper-with-me

Papers

360BEV: Panoramic Semantic Mapping for Indoor Bird's-Eye View

2023-03-21 · Zhifeng Teng, Jiaming Zhang, Kailun Yang, Kunyu Peng, Hao Shi, Simon Reiß, Ke Cao, Rainer Stiefelhagen

Seeing only a tiny part of the whole is not knowing the full circumstance. Bird's-eye-view (BEV) perception, a process of obtaining allocentric maps from egocentric views, is restricted when using a narrow Field of View (FoV) alone. In this work, mapping from 360{\deg} panoramas to BEV semantics, the 360BEV task, is established for the first time to achieve holistic representations of indoor scenes in a top-down view. Instead of relying on narrow-FoV image sequences, a panoramic image with depth information is sufficient to generate a holistic BEV semantic map. To benchmark 360BEV, we present two indoor datasets, 360BEV-Matterport and 360BEV-Stanford, both of which include egocentric panoramic images and semantic segmentation labels, as well as allocentric semantic maps. Besides delving deep into different mapping paradigms, we propose a dedicated solution for panoramic semantic mapping, namely 360Mapper. Through extensive experiments, our methods achieve 44.32% and 45.78% in mIoU on both datasets respectively, surpassing previous counterparts with gains of +7.60% and +9.70% in mIoU. Code and datasets are available at the project page: https://jamycheung.github.io/360BEV.html.

📄 PDF Abstract BibTeX arXiv:2303.11910

Code (1)

jamycheung/360BEV 공식 구현 pytorch

Tasks

Semantic Segmentation

Similar Papers 제목 키워드 기반

OneBEV: Using One Panoramic Image for Bird's-Eye-View Semantic Mapping

2024-09-20 · Jiale Wei, Junwei Zheng, Ruiping Liu, Jie Hu 외

In the field of autonomous driving, Bird's-Eye-View (BEV) perception has attracted increasing attention in the community since it provides more comprehensive information compared with pinhole front-view images and panora…

Autonomous DrivingMamba

GEM-Occ: From Visual Geometry Evidence to Embodied Semantic Occupancy Memory

2026-07-06 · Hu Zhu, Bohan Li, Xianda Guo, Hongsi Liu 외 arxiv

Semantic occupancy provides a structured spatial memory for embodied indoor agents by jointly representing occupied regions, observed free space, unknown areas, and object semantics. However, existing indoor occupancy be…

Argus: Metric Panoramic 3D Reconstruction for Indoor Scenes

2026-06-29 · Xi Li, Linyuan Li, Yan Wu, Tong Rao 외 arxiv

Metric feed-forward 3D reconstruction for panoramic data remains under-explored due to the lack of large-scale panoramic RGB-D training data. We present Realsee3D, a hybrid dataset of 10K indoor scenes (1K real, 9K synth…

Camera Pose EstimationMulti-Task Learning3D ReconstructionDepth Estimation

Bird's-Eye-View Scene Graph for Vision-Language Navigation

2023-08-09 · ICCV 2023 1 · Rui Liu, Xiaohan Wang, Wenguan Wang, Yi Yang

Vision-language navigation (VLN), which entails an agent to navigate 3D environments following human instructions, has shown great advances. However, current agents are built upon panoramic observations, which hinders th…

NavigateVision-Language Navigation

Matterport3D: Learning from RGB-D Data in Indoor Environments

2017-09-18 · Angel Chang, Angela Dai, Thomas Funkhouser, Maciej Halber 외

Access to large, diverse RGB-D datasets is critical for training RGB-D scene understanding algorithms. However, existing datasets still cover only a limited number of views or a restricted scale of spaces. In this paper,…

General ClassificationScene UnderstandingSemantic Segmentation