paper-with-me

홈 › Papers

F?D: On understanding the role of deep feature spaces on face generation evaluation

2023-05-31 · Krish Kabra, Guha Balakrishnan

Perceptual metrics, like the Fr\'echet Inception Distance (FID), are widely used to assess the similarity between synthetically generated and ground truth (real) images. The key idea behind these metrics is to compute errors in a deep feature space that captures perceptually and semantically rich image features. Despite their popularity, the effect that different deep features and their design choices have on a perceptual metric has not been well studied. In this work, we perform a causal analysis linking differences in semantic attributes and distortions between face image distributions to Fr\'echet distances (FD) using several popular deep feature spaces. A key component of our analysis is the creation of synthetic counterfactual faces using deep face generators. Our experiments show that the FD is heavily influenced by its feature space's training dataset and objective function. For example, FD using features extracted from ImageNet-trained models heavily emphasize hats over regions like the eyes and mouth. Moreover, FD using features from a face gender classifier emphasize hair length more than distances in an identity (recognition) feature space. Finally, we evaluate several popular face generation models across feature spaces and find that StyleGAN2 consistently ranks higher than other face generators, except with respect to identity (recognition) features. This suggests the need for considering multiple feature spaces when evaluating generative models and using feature spaces that are tuned to nuances of the domain of interest.

📄 PDF Abstract BibTeX arXiv:2305.20048

Code (0)

등록된 구현이 없습니다.

Tasks

counterfactualFace Generation

Methods 이 논문이 사용한 방법론

HuMan(Expedia)||How do I get a human at Expedia? How do I get a human at Expedia? How Do I Get a Human at Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Real-Time Help & Exclusive…
R1 Regularization R_INLINE_MATH_1 Regularization is a regularization technique and gradient penalty for training [generative adversarial…
Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…
Path Length Regularization 설명 없음
Weight Demodulation 설명 없음

Similar Papers 제목 키워드 기반

Uncovering Bias in Face Generation Models

2023-02-22 · Cristian Muñoz, Sara Zannone, Umar Mohammed, Adriano Koshiyama

Recent advancements in GANs and diffusion models have enabled the creation of high-resolution, hyper-realistic images. However, these models may misrepresent certain social groups and present bias. Understanding bias in …

AttributeDecision MakingFace GenerationFairness

Understanding In-Context Learning from Repetitions

2023-09-30 · Jianhao Yan, Jin Xu, Chiyu Song, Chenming Wu 외

This paper explores the elusive mechanism underpinning in-context learning in Large Language Models (LLMs). Our work provides a novel perspective by examining in-context learning via the lens of surface repetitions. We q…

In-Context LearningText Generation

Detecting Parking Spaces in a Parcel using Satellite Images

2019-08-28 · Murugesan Vadivel, Selvakumar Murugan, Suriyadeepan Ramamoorthy, Vaidheeswaran Archana 외

Remote Sensing Images from satellites have been used in various domains for detecting and understanding structures on the ground surface. In this work, satellite images were used for localizing parking spaces and vehicle…

RULLS: Randomized Union of Locally Linear Subspaces for Feature Engineering

2018-04-25 · Namita Lokare, Jorge Silva, Ilknur Kaynar Kabul

Feature engineering plays an important role in the success of a machine learning model. Most of the effort in training a model goes into data preparation and choosing the right representation. In this paper, we propose a…

ClusteringFeature EngineeringGeneral Classification

OmniBind: Large-scale Omni Multimodal Representation via Binding Spaces

2024-07-16 · Zehan Wang, Ziang Zhang, Hang Zhang, Luping Liu 외

Recently, human-computer interaction with various modalities has shown promising applications, like GPT-4o and Gemini. Given the foundational role of multimodal joint representation in understanding and generation pipeli…