Boosting Detection in Crowd Analysis via Underutilized Output Features
Detection-based methods have been viewed unfavorably in crowd analysis due to their poor performance in dense crowds. However, we argue that the potential of these methods has been underestimated, as they offer crucial information for crowd analysis that is often ignored. Specifically, the area size and confidence score of output proposals and bounding boxes provide insight into the scale and density of the crowd. To leverage these underutilized features, we propose Crowd Hat, a plug-and-play module that can be easily integrated with existing detection models. This module uses a mixed 2D-1D compression technique to refine the output features and obtain the spatial and numerical distribution of crowd-specific information. Based on these features, we further propose region-adaptive NMS thresholds and a decouple-then-align paradigm that address the major limitations of detection-based methods. Our extensive evaluations on various crowd analysis tasks, including crowd counting, localization, and detection, demonstrate the effectiveness of utilizing output features and the potential of detection-based methods in crowd analysis.
Code (1)
Tasks
Crowd CountingSimilar Papers 제목 키워드 기반
Cloud-based Federated Boosting for Mobile Crowdsensing
The application of federated extreme gradient boosting to mobile crowdsensing apps brings several benefits, in particular high performance on efficiency and classification. However, it also brings a new challenge for dat…
Federated LearningGeneral ClassificationGenerative Adversarial NetworkPrivacy Preserving+1$L_2$Boosting for Economic Applications
In the recent years more and more high-dimensional data sets, where the number of parameters $p$ is high compared to the number of observations $n$ or even larger, are available for applied researchers. Boosting algorith…
AttributeVisibility Guided NMS: Efficient Boosting of Amodal Object Detection in Crowded Traffic Scenes
Object detection is an important task in environment perception for autonomous driving. Modern 2D object detection frameworks such as Yolo, SSD or Faster R-CNN predict multiple bounding boxes per object that are refined …
2D Object DetectionAutonomous DrivingObjectobject-detection+1Towards Better User Studies in Computer Graphics and Vision
Online crowdsourcing platforms have made it increasingly easy to perform evaluations of algorithm outputs with survey questions like "which image is better, A or B?", leading to their proliferation in vision and graphics…
Using Depth for Pixel-Wise Detection of Adversarial Attacks in Crowd Counting
State-of-the-art methods for counting people in crowded scenes rely on deep networks to estimate crowd density. While effective, deep learning approaches are vulnerable to adversarial attacks, which, in a crowd-counting …
Crowd CountingDensity Estimation