paper-with-me

홈 › Papers

Embodied Learning for Lifelong Visual Perception

2021-12-28 · David Nilsson, Aleksis Pirinen, Erik Gärtner, Cristian Sminchisescu

We study lifelong visual perception in an embodied setup, where we develop new models and compare various agents that navigate in buildings and occasionally request annotations which, in turn, are used to refine their visual perception capabilities. The purpose of the agents is to recognize objects and other semantic classes in the whole building at the end of a process that combines exploration and active visual learning. As we study this task in a lifelong learning context, the agents should use knowledge gained in earlier visited environments in order to guide their exploration and active learning strategy in successively visited buildings. We use the semantic segmentation performance as a proxy for general visual perception and study this novel task for several exploration and annotation methods, ranging from frontier exploration baselines which use heuristic active learning, to a fully learnable approach. For the latter, we introduce a deep reinforcement learning (RL) based agent which jointly learns both navigation and active learning. A point goal navigation formulation, coupled with a global planner which supplies goals, is integrated into the RL model in order to provide further incentives for systematic exploration of novel scenes. By performing extensive experiments on the Matterport3D dataset, we show how the proposed agents can utilize knowledge from previously explored scenes when exploring new ones, e.g. through less granular exploration and less frequent requests for annotations. The results also suggest that a learning-based agent is able to use its prior visual knowledge more effectively than heuristic alternatives.

📄 PDF Abstract BibTeX arXiv:2112.14084

Code (0)

등록된 구현이 없습니다.

Tasks

Active LearningDeep Reinforcement LearningLifelong learningNavigateReinforcement Learning (RL)Semantic Segmentation

Similar Papers 제목 키워드 기반

The Empirical Impact of Forgetting and Transfer in Continual Visual Odometry

2024-06-03 · Paolo Cudrano, Xiaoyu Luo, Matteo Matteucci

As robotics continues to advance, the need for adaptive and continuously-learning embodied agents increases, particularly in the realm of assistance robotics. Quick adaptability and long-term information retention are es…

Lifelong learningTransfer LearningVisual Odometry

Ella: Embodied Social Agents with Lifelong Memory

2025-06-30 · Hongxin Zhang, Zheyuan Zhang, Zeyuan Wang, Zunzhe Zhang 외

We introduce Ella, an embodied social agent capable of lifelong learning within a community in a 3D open world, where agents accumulate experiences and acquire knowledge through everyday visual observations and social in…

Lifelong learning

Lifelong Embodied Navigation Learning

2026-03-06 · Xudong Wang, Jiahua Dong, Baichen Liu, Qi Lyu 외 arxiv

Embodied navigation agents powered by large language models have shown strong performance on individual tasks but struggle to continually acquire new navigation skills, which suffer from catastrophic forgetting. We forma…

3D-Mem: 3D Scene Memory for Embodied Exploration and Reasoning

2024-11-23 · CVPR 2025 1 · Yuncong Yang, Han Yang, Jiachen Zhou, Peihao Chen 외

Constructing compact and informative 3D scene representations is essential for effective embodied exploration and reasoning, especially in complex environments over extended periods. Existing representations, such as obj…

Management

AllDayNav: Lifelong Navigation via Real-World Reinforcement Learning

2026-06-09 · Hang Yin, Yinan Liang, Jiazhao Zhang, Jiahang Liu 외 arxiv

Lifelong embodied navigation in dynamic environments requires robots to form persistent scene understanding from fragmentary observations, which remains difficult for existing methods that rely on explicit maps or scene …

Reinforcement LearningScene Understanding