paper-with-me

Papers

A Standardized Framework For Evaluating Gene Expression Generative Models

2026-03-11 · Andrea Rubbi, Andrea Giuseppe Di Francesco, Mohammad Lotfollahi, Pietro Liò arxiv

The rapid development of generative models for single-cell gene expression data has created an urgent need for standardised evaluation frameworks. Current evaluation practices suffer from inconsistent metric implementations, incomparable hyperparameter choices, and a lack of biologically-grounded metrics. We present Generated Genetic Expression Evaluator (GGE), an open-source Python framework that addresses these challenges by providing a comprehensive suite of distributional metrics with explicit computation space options and biologically-motivated evaluation through differentially expressed gene (DEG)-focused analysis and perturbation-effect correlation, enabling standardized reporting and reproducible benchmarking. Through extensive analysis of the single-cell generative modeling literature, we identify that no standardized evaluation protocol exists. Methods report incomparable metrics computed in different spaces with different hyperparameters. We demonstrate that metric values vary substantially depending on implementation choices, highlighting the critical need for standardization. GGE enables fair comparison across generative approaches and accelerates progress in perturbation response prediction, cellular identity modeling, and counterfactual inference.

📄 PDF Abstract BibTeX arXiv:2603.11244

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

CGCE: A Chinese Generative Chat Evaluation Benchmark for General and Financial Domains

2023-05-23 · Xuanyu Zhang, Bingbing Li, Qing Yang

Generative chat models, such as ChatGPT and GPT-4, have revolutionized natural language generation (NLG) by incorporating instructions and human feedback to achieve significant performance improvements. However, the lack…

Text Generation

A novel database of Children's Spontaneous Facial Expressions (LIRIS-CSE)

2018-12-04 · Rizwan Ahmed Khan, Crenn Arthur, Alexandre Meyer, Saida Bouakaz

Computing environment is moving towards human-centered designs instead of computer centered designs and human's tend to communicate wealth of information through affective states or expressions. Traditional Human Compute…

BenchmarkingFacial Expression RecognitionFacial Expression Recognition (FER)Transfer Learning

A Framework for Human Evaluation of Large Language Models in Healthcare Derived from Literature Review

2024-05-04 · Thomas Yu CHow Tam, Sonish Sivarajkumar, Sumit Kapoor, Alisa V Stolyar 외

With generative artificial intelligence (AI), particularly large language models (LLMs), continuing to make inroads in healthcare, it is critical to supplement traditional automated evaluations with human evaluations. Un…

HYPE-C: Evaluating Image Completion Models Through Standardized Crowdsourcing

2021-01-01 · Emily Walters, Weifeng Chen, Jia Deng

A significant obstacle to the development of new image completion models is the lack of a standardized evaluation metric that reflects human judgement. Recent work has proposed the use of human evaluation for image synth…

Image Generation

Computational Hermeneutics: Evaluating generative AI as a cultural technology

2026-03-31 · Cody Kommers, Ruth Ahnert, Maria Antoniak, Emmanouil Benetos 외 arxiv

Generative AI systems are increasingly recognized as cultural technologies, yet current evaluation frameworks often treat culture as a variable to be measured rather than fundamental to the system's operation. Drawing on…