paper-with-me

홈 › Papers

AID: Pushing the Performance Boundary of Human Pose Estimation with Information Dropping Augmentation

2020-08-17 · Junjie Huang, Zheng Zhu, Guan Huang, Dalong Du

Both appearance cue and constraint cue are vital for human pose estimation. However, there is a tendency in most existing works to overfitting the former and overlook the latter. In this paper, we propose Augmentation by Information Dropping (AID) to verify and tackle this dilemma. Alone with AID as a prerequisite for effectively exploiting its potential, we propose customized training schedules, which are designed by analyzing the pattern of loss and performance in training process from the perspective of information supplying. In experiments, as a model-agnostic approach, AID promotes various state-of-the-art methods in both bottom-up and top-down paradigms with different input sizes, frameworks, backbones, training and testing sets. On popular COCO human pose estimation test set, AID consistently boosts the performance of different configurations by around 0.6 AP in top-down paradigm and up to 1.5 AP in bottom-up paradigm. On more challenging CrowdPose dataset, the improvement is more than 1.5 AP. As AID successfully pushes the performance boundary of human pose estimation problem by considerable margin and sets a new state-of-the-art, we hope AID to be a regular configuration for training human pose estimators. The source code will be publicly available for further research.

📄 PDF Abstract BibTeX arXiv:2008.07139

Code (2)

HuangJunJie2017/UDP-Pose 공식 구현 mxnet
open-mmlab/mmpose pytorch

Tasks

Multi-Person Pose EstimationPose Estimation

Methods 이 논문이 사용한 방법론

Batch Normalization 설명 없음
Residual Connection 설명 없음
Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…
ReLU How Do I Communicate to Expedia? How Do I Communicate to Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Live Support & Special Travel…
HRNet HRNet, or High-Resolution Net, is a general purpose convolutional neural network for tasks like semantic segmentation, object detection and image classification. It is…

Similar Papers 제목 키워드 기반

2D Human Pose Estimation: A Survey

2022-04-15 · Haoming Chen, Runyang Feng, Sifan Wu, Hao Xu 외

Human pose estimation aims at localizing human anatomical keypoints or body parts in the input data (e.g., images, videos, or signals). It forms a crucial component in enabling machines to have an insightful understandin…

2D Human Pose EstimationKeypoint DetectionPose EstimationSurvey

PnPNet: Pull-and-Push Networks for Volumetric Segmentation with Boundary Confusion

2023-12-13 · Xin You, Ming Ding, Minghui Zhang, Hanxiao Zhang 외

Precise boundary segmentation of volumetric images is a critical task for image-guided diagnosis and computer-assisted intervention, especially for boundary confusion in clinical practice. However, U-shape networks canno…

Pushing the Boundaries of Boundary Detection using Deep Learning

2015-11-23 · Iasonas Kokkinos

In this work we show that adapting Deep Convolutional Neural Network training to the task of boundary detection can result in substantial improvements over the current state-of-the-art in boundary detection. Our contri…

Boundary DetectionDeep LearningSegmentationSemantic Segmentation

Conditional Boundary Loss for Semantic Segmentation

2023-07-05 · IEEE Transactions on Image Processing 2023 7 · Dongyue Wu, Zilin Guo, Aoyan Li, Changqian Yu 외

Improving boundary segmentation results has recently attracted increasing attention in the field of semantic segmentation. Since existing popular methods usually exploit the long-range context, the boundary cues are obsc…

SegmentationSemantic Segmentation

Stationary Point Losses for Robust Model

2023-02-19 · Weiwei Gao, Dazhi Zhang, Yao Li, Zhichang Guo 외

The inability to guarantee robustness is one of the major obstacles to the application of deep learning models in security-demanding domains. We identify that the most commonly used cross-entropy (CE) loss does not guara…

model