paper-with-me

홈 › Papers

Identity From Here, Pose From There: Self-Supervised Disentanglement and Generation of Objects Using Unlabeled Videos

2019-10-01 · ICCV 2019 10 · Fanyi Xiao, Haotian Liu, Yong Jae Lee

We propose a novel approach that disentangles the identity and pose of objects for image generation. Our model takes as input an ID image and a pose image, and generates an output image with the identity of the ID image and the pose of the pose image. Unlike most previous unsupervised work which rely on cyclic constraints, which can often be brittle, we instead propose to learn this in a self-supervised way. Specifically, we leverage unlabeled videos to automatically construct pseudo ground-truth targets to directly supervise our model. To enforce disentanglement, we propose a novel disentanglement loss, and to improve realism, we propose a pixel-verification loss in which the generated image's pixels must trace back to the ID input. We conduct extensive experiments on both synthetic and real images to demonstrate improved realism, diversity, and ID/pose disentanglement compared to existing methods.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

DisentanglementDiversityImage Generation

Similar Papers 제목 키워드 기반

Identity-Disentangled Adversarial Augmentation for Self-supervised Learning

2021-09-29 · Kaiwen Yang, Tianyi Zhou, Xinmei Tian, DaCheng Tao

Data augmentation is critical to contrastive self-supervised learning, whose goal is to distinguish a sample's augmentations (positives) from other samples (negatives). However, strong augmentations may change the sample…

Contrastive LearningData AugmentationSelf-Supervised Learning

Mandarin-English Code-switching Speech Recognition with Self-supervised Speech Representation Models

2021-10-07 · Liang-Hsuan Tseng, Yu-Kuan Fu, Heng-Jui Chang, Hung-Yi Lee

Code-switching (CS) is common in daily conversations where more than one language is used within a sentence. The difficulties of CS speech recognition lie in alternating languages and the lack of transcribed data. Theref…

Language IdentificationSelf-Supervised LearningSentencespeech-recognition+1

Asymmetric Mask Scheme for Self-Supervised Real Image Denoising

2024-07-09 · Xiangyu Liao, Tianheng Zheng, Jiayu Zhong, Pingping Zhang 외

In recent years, self-supervised denoising methods have gained significant success and become critically important in the field of image restoration. Among them, the blind spot network based methods are the most typical …

DenoisingImage DenoisingImage Restoration

Intra-Camera Supervised Person Re-Identification: A New Benchmark

2019-08-27 · Xiangping Zhu, Xiatian Zhu, Minxian Li, Vittorio Murino 외

Existing person re-identification (re-id) methods rely mostly on a large set of inter-camera identity labelled training data, requiring a tedious data collection and annotation process therefore leading to poor scalabili…

Multi-Label LearningPerson Re-Identification

Self-supervised Data Bootstrapping for Deep Optical Character Recognition of Identity Documents

2019-08-12 · Oliver Mothes, Joachim Denzler

The essential task of verifying person identities at airports and national borders is very time consuming. To accelerate it, optical character recognition for identity documents (IDs) using dictionaries is not appropriat…

Optical Character RecognitionOptical Character Recognition (OCR)