paper-with-me

Papers

Trade-offs in Image Generation: How Do Different Dimensions Interact?

2025-07-29 · Sicheng Zhang, Binzhu Xie, Zhonghao Yan, Yuli Zhang, Donghao Zhou, Xiaofei Chen, Shi Qiu, Jiaqi Liu, Guoyang Xie, Zhichao Lu arxiv

Model performance in text-to-image (T2I) and image-to-image (I2I) generation often depends on multiple aspects, including quality, alignment, diversity, and robustness. However, models' complex trade-offs among these dimensions have rarely been explored due to (1) the lack of datasets that allow fine-grained quantification of these trade-offs, and (2) the use of a single metric for multiple dimensions. To bridge this gap, we introduce TRIG-Bench (Trade-offs in Image Generation), which spans 10 dimensions (Realism, Originality, Aesthetics, Content, Relation, Style, Knowledge, Ambiguity, Toxicity, and Bias), contains 40,200 samples, and covers 132 pairwise dimensional subsets. Furthermore, we develop TRIGScore, a VLM-as-judge metric that automatically adapts to various dimensions. Based on TRIG-Bench and TRIGScore, we evaluate 14 models across T2I and I2I tasks. In addition, we propose the Relation Recognition System to generate the Dimension Trade-off Map (DTM) that visualizes the trade-offs among model-specific capabilities. Our experiments demonstrate that DTM consistently provides a comprehensive understanding of the trade-offs between dimensions for each type of generative model. Notably, we show that the model's dimension-specific weaknesses can be mitigated through fine-tuning on DTM to enhance overall performance. Code is available at: https://github.com/fesvhtr/TRIG

📄 PDF Abstract BibTeX arXiv:2507.22100

Code (0)

등록된 구현이 없습니다.

Tasks

Image Generation

Similar Papers 제목 키워드 기반

Sample, computation vs storage tradeoffs for classification using tensor subspace models

2017-06-18 · Mohammadhossein Chaghazardi, Shuchin Aeron

In this paper, we exhibit the tradeoffs between the (training) sample, computation and storage complexity for the problem of supervised classification using signal subspace estimation. Our main tool is the use of tensor …

General Classification

Can Large Language Models Make Everyone Happy?

2026-02-11 · Usman Naseem, Gautam Siddharth Kashyap, Ebad Shabbir, Sushant Kumar Ray 외 arxiv

Misalignment in Large Language Models (LLMs) refers to the failure to simultaneously satisfy safety, value, and cultural dimensions, leading to behaviors that diverge from human expectations in real-world settings where …

The influence of the composition of tradeoffs on the generation of differentiated cells

2016-08-30

We study the emergence of cell differentiation under the assumption of the existence of a given number of tradeoffs between genes encoding different functions. In the model the viability of colonies is determined by the …

No Free Lunch for Synthetic Images under Data Scarcity Conditions

2026-06-01 · Borja Arroyo Galende, Alejandro Almodóvar, Patricia A. Apellániz, Juan Parras 외 arxiv

This study investigates the trade-offs between fidelity, privacy, and utility in synthetic data generation under conditions of data scarcity and privacy sensitivity. We propose an evaluation framework that jointly assess…

Synthetic Data Generation

Performance-Efficiency Trade-offs in Unsupervised Pre-training for Speech Recognition

2021-09-14 · Felix Wu, Kwangyoun Kim, Jing Pan, Kyu Han 외

This paper is a study of performance-efficiency trade-offs in pre-trained models for automatic speech recognition (ASR). We focus on wav2vec 2.0, and formalize several architecture designs that influence both the model p…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition+1