paper-with-me

홈 › Papers

SegmATRon: Embodied Adaptive Semantic Segmentation for Indoor Environment

2023-10-18 · Tatiana Zemskova, Margarita Kichik, Dmitry Yudin, Aleksei Staroverov, Aleksandr Panov

This paper presents an adaptive transformer model named SegmATRon for embodied image semantic segmentation. Its distinctive feature is the adaptation of model weights during inference on several images using a hybrid multicomponent loss function. We studied this model on datasets collected in the photorealistic Habitat and the synthetic AI2-THOR Simulators. We showed that obtaining additional images using the agent's actions in an indoor environment can improve the quality of semantic segmentation. The code of the proposed approach and datasets are publicly available at https://github.com/wingrune/SegmATRon.

📄 PDF Abstract BibTeX arXiv:2310.12031

Code (1)

wingrune/segmatron 공식 구현 pytorch

Tasks

SegmentationSemantic Segmentation

Similar Papers 제목 키워드 기반

The Replica Dataset: A Digital Replica of Indoor Spaces

2019-06-13 · Julian Straub, Thomas Whelan, Lingni Ma, Yufan Chen 외

We introduce Replica, a dataset of 18 highly photo-realistic 3D indoor scene reconstructions at room and building scale. Each scene consists of a dense mesh, high-resolution high-dynamic-range (HDR) textures, per-primiti…

3D Scene ReconstructionInstruction FollowingQuestion AnsweringSemantic Segmentation

JSMNet Improving Indoor Point Cloud Semantic and Instance Segmentation through Self-Attention and Multiscale

2023-09-14 · Shuochen Xu, Zhenxin Zhang

The semantic understanding of indoor 3D point cloud data is crucial for a range of subsequent applications, including indoor service robots, navigation systems, and digital twin engineering. Global features are crucial f…

Instance SegmentationSegmentationSemantic Segmentation

Semantic Mapping in Indoor Embodied AI -- A Survey on Advances, Challenges, and Future Directions

2025-01-10 · Sonia Raychaudhuri, Angel X. Chang

Intelligent embodied agents (e.g. robots) need to perform complex semantic tasks in unfamiliar environments. Among many skills that the agents need to possess, building and maintaining a semantic map of the environment i…

Embodied Amodal Recognition: Learning to Move to Perceive Objects

2019-10-01 · ICCV 2019 10 · Jianwei Yang, Zhile Ren, Mingze Xu, Xinlei Chen 외

Passive visual systems typically fail to recognize objects in the amodal setting where they are heavily occluded. In contrast, humans and other embodied agents have the ability to move in the environment and actively con…

ObjectObject LocalizationSemantic Segmentation

Embodied Visual Active Learning for Semantic Segmentation

2020-12-17 · David Nilsson, Aleksis Pirinen, Erik Gärtner, Cristian Sminchisescu

We study the task of embodied visual active learning, where an agent is set to explore a 3d environment with the goal to acquire visual scene understanding by actively selecting views for which to request annotation. Whi…

Active LearningDeep Reinforcement LearningScene UnderstandingSemantic Segmentation