paper-with-me

홈 › Papers

Active View Selection for Scene-level Multi-view Crowd Counting and Localization with Limited Labeling Budget

2025-09-20 · Qi Zhang, Bin Li, Antoni B. Chan, Hui Huang arxiv

Multi-view crowd counting and localization fuse the input multi-views for estimating the crowd number or locations on the ground. Existing methods mainly focus on accurately predicting on the crowd shown in the input views, which neglects the problem of choosing the `best' camera views to perceive all crowds well in the scene. Besides, existing view selection methods require massive labeled views and images, and lack the ability for cross-scene settings, reducing their application scenarios. Thus, in this paper, we study the view selection issue for better scene-level multi-view crowd counting and localization results with cross-scene ability and limited label demand, instead of input-view-level results. We first propose an independent view selection method (IVS) that considers view and scene geometries in the view selection strategy and conducts the view selection, labeling, and downstream tasks independently. Based on IVS, we also put forward an active view selection method (AVS) that jointly conducts the view selection, labeling, and downstream tasks. In AVS, we actively select the labeled views and consider both the view/scene geometries and the predictions of the downstream task models in the view selection process. Experiments on multi-view counting and localization tasks demonstrate the cross-scene and the limited label demand advantages of the proposed active view selection method (AVS), outperforming existing methods and with wider application scenarios.

📄 PDF Abstract BibTeX arXiv:2509.16684

Code (0)

등록된 구현이 없습니다.

Tasks

Crowd Counting

Similar Papers 제목 키워드 기반

Improving Viewpoint-Independent Object-Centric Representations through Active Viewpoint Selection

2024-11-01 · Yinxuan Huang, Chengmin Gao, Bin Li, xiangyang xue

Given the complexities inherent in visual scenes, such as object occlusion, a comprehensive understanding often requires observation from multiple viewpoints. Existing multi-viewpoint object-centric learning methods typi…

Object

Domes to Drones: Self-Supervised Active Triangulation for 3D Human Pose Reconstruction

2019-12-01 · NeurIPS 2019 12 · Aleksis Pirinen, Erik Gärtner, Cristian Sminchisescu

Existing state-of-the-art estimation systems can detect 2d poses of multiple people in images quite reliably. In contrast, 3d pose estimation from a single image is ill-posed due to occlusion and depth ambiguities. Assum…

2D Pose Estimation3D Pose Estimation3D ReconstructionDeep Reinforcement Learning+2

Active Neural 3D Reconstruction with Colorized Surface Voxel-based View Selection

2024-05-04 · Hyunseo Kim, Hyeonseo Yang, Taekyung Kim, Yoonsung Kim 외

Active view selection in 3D scene reconstruction has been widely studied since training on informative views is critical for reconstruction. Recently, Neural Radiance Fields (NeRF) variants have shown promising results i…

3D Reconstruction3D Scene ReconstructionActive 3D ReconstructionActive Learning+2

Exploring Active Learning for Label-Efficient Training of Semantic Neural Radiance Field

2025-07-23 · Yuzhe Zhu, Lile Cai, Kangkang Lu, Fayao Liu 외 arxiv

Neural Radiance Field (NeRF) models are implicit neural scene representation methods that offer unprecedented capabilities in novel view synthesis. Semantically-aware NeRFs not only capture the shape and radiance of a sc…

Novel View SynthesisActive Learning

Uncertainty-Driven Active Vision for Implicit Scene Reconstruction

2022-10-03 · Edward J. Smith, Michal Drozdzal, Derek Nowrouzezahrai, David Meger 외

Multi-view implicit scene reconstruction methods have become increasingly popular due to their ability to represent complex scene details. Recent efforts have been devoted to improving the representation of input informa…

Scene Understanding