paper-with-me

Papers

Unsupervised Cross-Modal Alignment for Multi-Person 3D Pose Estimation

2020-08-04 · ECCV 2020 8 · Jogendra Nath Kundu, Ambareesh Revanur, Govind Vitthal Waghmare, Rahul Mysore Venkatesh, R. Venkatesh Babu

We present a deployment friendly, fast bottom-up framework for multi-person 3D human pose estimation. We adopt a novel neural representation of multi-person 3D pose which unifies the position of person instances with their corresponding 3D pose representation. This is realized by learning a generative pose embedding which not only ensures plausible 3D pose predictions, but also eliminates the usual keypoint grouping operation as employed in prior bottom-up approaches. Further, we propose a practical deployment paradigm where paired 2D or 3D pose annotations are unavailable. In the absence of any paired supervision, we leverage a frozen network, as a teacher model, which is trained on an auxiliary task of multi-person 2D pose estimation. We cast the learning as a cross-modal alignment problem and propose training objectives to realize a shared latent space between two diverse modalities. We aim to enhance the model's ability to perform beyond the limiting teacher network by enriching the latent-to-3D pose mapping using artificially synthesized multi-person 3D scene samples. Our approach not only generalizes to in-the-wild images, but also yields a superior trade-off between speed and performance, compared to prior top-down approaches. Our approach also yields state-of-the-art multi-person 3D pose estimation performance among the bottom-up approaches under consistent supervision levels.

📄 PDF Abstract BibTeX arXiv:2008.01388

Code (1)

revanurambareesh/multiperson tf

Tasks

2D Pose Estimation3D Human Pose Estimation3D Multi-Person Pose Estimation3D Pose Estimationcross-modal alignmentPose EstimationUnsupervised 3D Human Pose EstimationUnsupervised 3D Multi-Person Pose Estimation

Methods 이 논문이 사용한 방법론

Dogecoin Customer Service Number +1-833-534-1729 설명 없음

Similar Papers 제목 키워드 기반

Unsupervised Visible-Infrared ReID via Pseudo-label Correction and Modality-level Alignment

2024-04-10 · Yexin Liu, Weiming Zhang, Athanasios V. Vasilakos, Lin Wang

Unsupervised visible-infrared person re-identification (UVI-ReID) has recently gained great attention due to its potential for enhancing human detection in diverse environments without labeling. Previous methods utilize …

ClusteringContrastive LearningCross-Modality Person Re-identificationHuman Detection+2

Multi-Memory Matching for Unsupervised Visible-Infrared Person Re-Identification

2024-01-12 · Jiangming Shi, Xiangbo Yin, Yeyun Chen, Yachao Zhang 외

Unsupervised visible-infrared person re-identification (USL-VI-ReID) is a promising yet challenging retrieval task. The key challenges in USL-VI-ReID are to effectively generate pseudo-labels and establish pseudo-label c…

ClusteringPerson Re-IdentificationPseudo Label

FedEPA: Enhancing Personalization and Modality Alignment in Multimodal Federated Learning

2025-04-16 · Yu Zhang, Qingfeng Du, Jiaqi Lv

Federated Learning (FL) enables decentralized model training across multiple parties while preserving privacy. However, most FL systems assume clients hold only unimodal data, limiting their real-world applicability, as …

Contrastive LearningDiversityFederated Learning

Temporal Prototyping and Hierarchical Alignment for Unsupervised Video-based Visible-Infrared Person Re-Identification

2026-04-23 · Zhiyong Li, Wei Jiang, Haojie Liu, Mingyu Wang 외 arxiv

Visible-infrared person re-identification (VI-ReID) enables cross-modality identity matching for all-day surveillance, yet existing methods predominantly focus on the image level or rely heavily on costly identity annota…

Person Re-IdentificationContrastive Learning

Domain-Shared Learning and Gradual Alignment for Unsupervised Domain Adaptation Visible-Infrared Person Re-Identification

2025-11-20 · Nianchang Huang, Yi Xu, Ruida Xi, Ruida Xi 외 arxiv

Recently, Visible-Infrared person Re-Identification (VI-ReID) has achieved remarkable performance on public datasets. However, due to the discrepancies between public datasets and real-world data, most existing VI-ReID a…

Unsupervised Domain AdaptationPerson Re-Identification