Optimizing Camera Configurations for Multi-View Pedestrian Detection
Jointly considering multiple camera views (multi-view) is very effective for pedestrian detection under occlusion. For such multi-view systems, it is critical to have well-designed camera configurations, including camera locations, directions, and fields-of-view (FoVs). Usually, these configurations are crafted based on human experience or heuristics. In this work, we present a novel solution that features a transformer-based camera configuration generator. Using reinforcement learning, this generator autonomously explores vast combinations within the action space and searches for configurations that give the highest detection accuracy according to the training dataset. The generator learns advanced techniques like maximizing coverage, minimizing occlusion, and promoting collaboration. Across multiple simulation scenarios, the configurations generated by our transformer-based model consistently outperform random search, heuristic-based methods, and configurations designed by human experts, shedding light on future camera layout optimization.
Code (0)
등록된 구현이 없습니다.
Tasks
Pedestrian DetectionSimilar Papers 제목 키워드 기반
MV2GF: Multi-view Pedestrian Detection with a Visual Geometric Foundation Model
Multi-View Pedestrian Detection (MVPD) aims to detect pedestrians in the form of a bird's eye view map from multi-view images. Recent MVPD methods adopt a unified framework that projects 2D image features into a 3D world…
Pedestrian DetectionMVUDA: Unsupervised Domain Adaptation for Multi-view Pedestrian Detection
We address multi-view pedestrian detection in a setting where labeled data is collected using a multi-camera setup different from the one used for testing. While recent multi-view pedestrian detectors perform well on the…
Domain AdaptationPedestrian DetectionUnsupervised Domain AdaptationUnsupervised Multi-view Pedestrian Detection
With the prosperity of the video surveillance, multiple cameras have been applied to accurately locate pedestrians in a specific area. However, previous methods rely on the human-labeled annotations in every video frame …
Camera CalibrationPedestrian DetectionMulti-Camera Trajectory Forecasting: Pedestrian Trajectory Prediction in a Network of Cameras
We introduce the task of multi-camera trajectory forecasting (MCTF), where the future trajectory of an object is predicted in a network of cameras. Prior works consider forecasting trajectories in a single camera view. O…
Pedestrian Trajectory PredictionTrajectory ForecastingTrajectory PredictionKey Person Aided Re-identification in Partially Ordered Pedestrian Set
Ideally person re-identification seeks for perfect feature representation and metric model that re-identify all various pedestrians well in non-overlapping views at different locations with different camera configuration…
Person Re-Identification