paper-with-me

Papers

ActLoc: Learning to Localize on the Move via Active Viewpoint Selection

2025-08-28 · Jiajie Li, Boyang Sun, Luca Di Giammarino, Hermann Blum, Marc Pollefeys arxiv

Reliable localization is critical for robot navigation, yet most existing systems implicitly assume that all viewing directions at a location are equally informative. In practice, localization becomes unreliable when the robot observes unmapped, ambiguous, or uninformative regions. To address this, we present ActLoc, an active viewpoint-aware planning framework for enhancing localization accuracy for general robot navigation tasks. At its core, ActLoc employs a largescale trained attention-based model for viewpoint selection. The model encodes a metric map and the camera poses used during map construction, and predicts localization accuracy across yaw and pitch directions at arbitrary 3D locations. These per-point accuracy distributions are incorporated into a path planner, enabling the robot to actively select camera orientations that maximize localization robustness while respecting task and motion constraints. ActLoc achieves stateof-the-art results on single-viewpoint selection and generalizes effectively to fulltrajectory planning. Its modular design makes it readily applicable to diverse robot navigation and inspection tasks.

📄 PDF Abstract BibTeX arXiv:2508.20981

Code (0)

등록된 구현이 없습니다.

Tasks

Robot Navigation

Similar Papers 제목 키워드 기반

Improving Viewpoint-Independent Object-Centric Representations through Active Viewpoint Selection

2024-11-01 · Yinxuan Huang, Chengmin Gao, Bin Li, xiangyang xue

Given the complexities inherent in visual scenes, such as object occlusion, a comprehensive understanding often requires observation from multiple viewpoints. Existing multi-viewpoint object-centric learning methods typi…

Object

ActiveVLA: Injecting Active Perception into Vision-Language-Action Models for Precise 3D Robotic Manipulation

2026-01-13 · Zhenyang Liu, Yongchong Gu, Yikai Wang, Xiangyang Xue 외 arxiv

Recent advances in robot manipulation have leveraged pre-trained vision-language models (VLMs) and explored integrating 3D spatial signals into these models for effective action prediction, giving rise to the promising v…

Robot Manipulation

Gaussian Process Models for HRTF based Sound-Source Localization and Active-Learning

2015-02-11 · Yuancheng Luo, Dmitry N. Zotkin, Ramani Duraiswami

From a machine learning perspective, the human ability localize sounds can be modeled as a non-parametric and non-linear regression problem between binaural spectral features of sound received at the ears (input) and the…

Active LearningregressionSound Source Localization

ViewAL: Active Learning with Viewpoint Entropy for Semantic Segmentation

2019-11-26 · CVPR 2020 6 · Yawar Siddiqui, Julien Valentin, Matthias Nießner

We propose ViewAL, a novel active learning strategy for semantic segmentation that exploits viewpoint consistency in multi-view datasets. Our core idea is that inconsistencies in model predictions across viewpoints provi…

Active LearningSemantic SegmentationSuperpixels

A Hybrid Model-Based and Model-Free Framework for Active Multi-View Viewpoint Optimization in Sonar Target Recognition

2026-06-13 · Yongkyoon Park, Jane Shin arxiv

This paper presents a hybrid model-based and model-free framework for active multi-view target recognition using forward-looking sonar. A convolutional neural network (CNN) provides data-driven observation likelihoods, w…