paper-with-me

홈 › Papers

Self-supervised 3D Representation Learning of Dressed Humans from Social Media Videos

2021-03-04 · CVPR 2021 1 · Yasamin Jafarian, Hyun Soo Park

A key challenge of learning a visual representation for the 3D high fidelity geometry of dressed humans lies in the limited availability of the ground truth data (e.g., 3D scanned models), which results in the performance degradation of 3D human reconstruction when applying to real-world imagery. We address this challenge by leveraging a new data resource: a number of social media dance videos that span diverse appearance, clothing styles, performances, and identities. Each video depicts dynamic movements of the body and clothes of a single person while lacking the 3D ground truth geometry. To learn a visual representation from these videos, we present a new self-supervised learning method to use the local transformation that warps the predicted local geometry of the person from an image to that of another image at a different time instant. This allows self-supervision by enforcing a temporal coherence over the predictions. In addition, we jointly learn the depths along with the surface normals that are highly responsive to local texture, wrinkle, and shade by maximizing their geometric consistency. Our method is end-to-end trainable, resulting in high fidelity depth estimation that predicts fine geometry faithful to the input real image. We further provide a theoretical bound of self-supervised learning via an uncertainty analysis that characterizes the performance of the self-supervised learning without training. We demonstrate that our method outperforms the state-of-the-art human depth estimation and human shape recovery approaches on both real and rendered images.

📄 PDF Abstract BibTeX arXiv:2103.03319

Code (1)

yasaminjafarian/HDNet_TikTok 공식 구현 tf

Tasks

3D Human ReconstructionDepth EstimationRepresentation LearningSelf-Supervised Learning

Similar Papers 제목 키워드 기반

Learning a Directional Soft Lane Affordance Model for Road Scenes Using Self-Supervision

2020-02-17 · Robin Karlsson, Erik Sjoberg

Humans navigate complex environments in an organized yet flexible manner, adapting to the context and implicit social rules. Understanding these naturally learned patterns of behavior is essential for applications such a…

Autonomous VehiclesNavigate

BotSSCL: Social Bot Detection with Self-Supervised Contrastive Learning

2024-02-06 · Mohammad Majid Akhtar, Navid Shadman Bhuiyan, Rahat Masood, Muhammad Ikram 외

The detection of automated accounts, also known as "social bots", has been an increasingly important concern for online social networks (OSNs). While several methods have been proposed for detecting social bots, signific…

Contrastive Learning

Interactively Learning Social Media Representations Improves News Source Factuality Detection

2023-09-26 · Nikhil Mehta, Dan Goldwasser

The rise of social media has enabled the widespread propagation of fake news, text that is published with an intent to spread misinformation and sway beliefs. Rapidly detecting fake news, especially as new events arise, …

Misinformation

Evaluating Speaker Identity Coding in Self-supervised Models and Humans

2024-06-14 · Gasser Elbanna

Speaker identity plays a significant role in human communication and is being increasingly used in societal applications, many through advances in machine learning. Speaker identity perception is an essential cognitive p…

Speaker Identification

Smile Like You Mean It: Driving Animatronic Robotic Face with Learned Models

2021-05-26 · Boyuan Chen, Yuhang Hu, Lianfeng Li, Sara Cummings 외

Ability to generate intelligent and generalizable facial expressions is essential for building human-like social robots. At present, progress in this field is hindered by the fact that each facial expression needs to be …

Camera CalibrationSelf-Supervised Learning