Multi-task human analysis in still images: 2D/3D pose, depth map, and multi-part segmentation
While many individual tasks in the domain of human analysis have recently received an accuracy boost from deep learning approaches, multi-task learning has mostly been ignored due to a lack of data. New synthetic datasets are being released, filling this gap with synthetic generated data. In this work, we analyze four related human analysis tasks in still images in a multi-task scenario by leveraging such datasets. Specifically, we study the correlation of 2D/3D pose estimation, body part segmentation and full-body depth estimation. These tasks are learned via the well-known Stacked Hourglass module such that each of the task-specific streams shares information with the others. The main goal is to analyze how training together these four related tasks can benefit each individual task for a better generalization. Results on the newly released SURREAL dataset show that all four tasks benefit from the multi-task approach, but with different combinations of tasks: while combining all four tasks improves 2D pose estimation the most, 2D pose improves neither 3D pose nor full-body depth estimation. On the other hand 2D parts segmentation can benefit from 2D pose but not from 3D pose. In all cases, as expected, the maximum improvement is achieved on those human body parts that show more variability in terms of spatial distribution, appearance and shape, e.g. wrists and ankles.
Code (0)
등록된 구현이 없습니다.
Tasks
2D Pose Estimation3D Pose EstimationDepth EstimationMulti-Task LearningPose EstimationMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Multi-task Deep Learning for Real-Time 3D Human Pose Estimation and Action Recognition
Human pose estimation and action recognition are related tasks since both problems are strongly dependent on the human body representation and analysis. Nonetheless, most recent methods in the literature handle the two p…
3D Human Pose EstimationAction RecognitionPose EstimationSkeleton Based Action RecognitionHuman Pose Estimation from Depth Images via Inference Embedded Multi-task Learning
Human pose estimation (i.e., locating the body parts / joints of a person) is a fundamental problem in human-computer interaction and multimedia applications. Significant progress has been made based on the development o…
Multi-Task LearningPose EstimationPose PredictionExploiting CLIP for Zero-shot HOI Detection Requires Knowledge Distillation at Multiple Levels
In this paper, we investigate the task of zero-shot human-object interaction (HOI) detection, a novel paradigm for identifying HOIs without the need for task-specific annotations. To address this challenging task, we emp…
Human-Object Interaction DetectionKnowledge DistillationLanguage ModelingLanguage ModellingFuture Aspects in Human Action Recognition: Exploring Emerging Techniques and Ethical Influences
Visual-based human action recognition can be found in various application fields, e.g., surveillance systems, sports analytics, medical assistive technologies, or human-robot interaction frameworks, and it concerns the i…
Action RecognitionSports AnalyticsTemporal Action LocalizationImage Distillation for Safe Data Sharing in Histopathology
Histopathology can help clinicians make accurate diagnoses, determine disease prognosis, and plan appropriate treatment strategies. As deep learning techniques prove successful in the medical domain, the primary challeng…
Dataset DistillationPrognosis