paper-with-me

Papers

Aggregate channel features for multi-view face detection

2014-07-15 · Bin Yang, Junjie Yan, Zhen Lei, Stan Z. Li

Face detection has drawn much attention in recent decades since the seminal work by Viola and Jones. While many subsequences have improved the work with more powerful learning algorithms, the feature representation used for face detection still can't meet the demand for effectively and efficiently handling faces with large appearance variance in the wild. To solve this bottleneck, we borrow the concept of channel features to the face detection domain, which extends the image channel to diverse types like gradient magnitude and oriented gradient histograms and therefore encodes rich information in a simple form. We adopt a novel variant called aggregate channel features, make a full exploration of feature design, and discover a multi-scale version of features with better performance. To deal with poses of faces in the wild, we propose a multi-view detection approach featuring score re-ranking and detection adjustment. Following the learning pipelines in Viola-Jones framework, the multi-view face detector using aggregate channel features shows competitive performance against state-of-the-art algorithms on AFW and FDDB testsets, while runs at 42 FPS on VGA images.

📄 PDF Abstract BibTeX arXiv:1407.4023

Code (0)

등록된 구현이 없습니다.

Tasks

Face Detectionmulti-view detectionRe-Ranking

Similar Papers 제목 키워드 기반

Dual-View Pyramid Pooling in Deep Neural Networks for Improved Medical Image Classification and Confidence Calibration

2024-08-06 · Xiaoqing Zhang, Qiushi Nie, Zunjie Xiao, Jilu Zhao 외

Spatial pooling (SP) and cross-channel pooling (CCP) operators have been applied to aggregate spatial features and pixel-wise features from feature maps in deep neural networks (DNNs), respectively. Their main goal is to…

Classificationimage-classificationImage ClassificationMedical Image Classification

Virtual Multi-view Fusion for 3D Semantic Segmentation

2020-07-26 · ECCV 2020 8 · Abhijit Kundu, Xiaoqi Yin, Alireza Fathi, David Ross 외

Semantic segmentation of 3D meshes is an important problem for 3D scene understanding. In this paper we revisit the classic multiview representation of 3D meshes and study several techniques that make them effective for …

2D Semantic Segmentation3D Semantic SegmentationScene UnderstandingSegmentation+1

Predicting the Factuality of Reporting of News Media Using Observations About User Attention in Their YouTube Channels

2021-08-27 · RANLP 2021 9 · Krasimira Bozhanova, Yoan Dinkov, Ivan Koychev, Maria Castaldo 외

We propose a novel framework for predicting the factuality of reporting of news media outlets by studying the user attention cycles in their YouTube channels. In particular, we design a rich set of features derived from …

MACCIF-TDNN: Multi aspect aggregation of channel and context interdependence features in TDNN-based speaker verification

2021-07-07 · Fangyuan Wang, Zhigang Song, Hongchen Jiang, Bo Xu

Most of the recent state-of-the-art results for speaker verification are achieved by X-vector and its subsequent variants. In this paper, we propose a new network architecture which aggregates the channel and context int…

Speaker Verification

Multiview Aggregation for Learning Category-Specific Shape Reconstruction

2019-07-01 · NeurIPS 2019 12 · Srinath Sridhar, Davis Rempe, Julien Valentin, Sofien Bouaziz 외

We investigate the problem of learning category-specific 3D shape reconstruction from a variable number of RGB views of previously unobserved object instances. Most approaches for multiview shape reconstruction operate o…

3D Shape ReconstructionObject