paper-with-me

Papers

AITTI: Learning Adaptive Inclusive Token for Text-to-Image Generation

2024-06-18 · Xinyu Hou, Xiaoming Li, Chen Change Loy

Despite the high-quality results of text-to-image generation, stereotypical biases have been spotted in their generated contents, compromising the fairness of generative models. In this work, we propose to learn adaptive inclusive tokens to shift the attribute distribution of the final generative outputs. Unlike existing de-biasing approaches, our method requires neither explicit attribute specification nor prior knowledge of the bias distribution. Specifically, the core of our method is a lightweight adaptive mapping network, which can customize the inclusive tokens for the concepts to be de-biased, making the tokens generalizable to unseen concepts regardless of their original bias distributions. This is achieved by tuning the adaptive mapping network with a handful of balanced and inclusive samples using an anchor loss. Experimental results demonstrate that our method outperforms previous bias mitigation methods without attribute specification while preserving the alignment between generative results and text descriptions. Moreover, our method achieves comparable performance to models that require specific attributes or editing directions for generation. Extensive experiments showcase the effectiveness of our adaptive inclusive tokens in mitigating stereotypical bias in text-to-image generation. The code will be available at https://github.com/itsmag11/AITTI.

📄 PDF Abstract BibTeX arXiv:2406.12805

Code (1)

itsmag11/aitti 공식 구현

Tasks

AttributeFairnessImage GenerationText to Image GenerationText-to-Image Generation

Similar Papers 제목 키워드 기반

Less than one percent of words would be affected by gender-inclusive language in German press texts

2024-02-06 · Carolin Müller-Spitzer, Samira Ochs, Alexander Koplenig, Jan-Oliver Rüdiger 외

Research on gender and language is tightly knitted to social debates on gender equality and non-discriminatory language use. Psycholinguistic scholars have made significant contributions in this field. However, corpus-ba…

Inclusive Interactive Collisions for Multi-View Consistent Compositional 3D Generation

2026-06-23 · Chang Liu, Mingwen Shao, Xiang Lv, Xinyuan Chen 외 arxiv

Recent breakthroughs in 3D generation have advanced notably with the development of text-to-image diffusion model. However, existing methods remain two practical challenges: (1) They primarily generate single 3D object, …

Scene Generation3D Generation

Inclusive Speaker Verification with Adaptive thresholding

2021-11-10 · Navdeep Jain, Hongcheng Wang

While using a speaker verification (SV) based system in a commercial application, it is important that customers have an inclusive experience irrespective of their gender, age, or ethnicity. In this paper, we analyze the…

Speaker Verification

Chitrakshara: A Large Multilingual Multimodal Dataset for Indian languages

2026-03-06 · Shaharukh Khan, Ali Faraz, Abhinav Ravi, Mohd Nauman 외 arxiv

Multimodal research has predominantly focused on single-image reasoning, with limited exploration of multi-image scenarios. Recent models have sought to enhance multi-image understanding through large-scale pretraining o…

ITI-GEN: Inclusive Text-to-Image Generation

2023-09-11 · ICCV 2023 1 · Cheng Zhang, Xuanbai Chen, Siqi Chai, Chen Henry Wu 외

Text-to-image generative models often reflect the biases of the training data, leading to unequal representations of underrepresented groups. This study investigates inclusive text-to-image generative models that generat…

AttributeImage GenerationText to Image GenerationText-to-Image Generation