paper-with-me

홈 › Papers

Enrichment Score: a better quantitative metric for evaluating the enrichment capacity of molecular docking models

2022-10-19 · Ian Scott Knight, Slava Naprienko, John J. Irwin

The standard quantitative metric for evaluating enrichment capacity known as $\textit{LogAUC}$ depends on a cutoff parameter that controls what the minimum value of the log-scaled x-axis is. Unless this parameter is chosen carefully for a given ROC curve, one of the two following problems occurs: either (1) some fraction of the first inter-decoy intervals of the ROC curve are simply thrown away and do not contribute to the metric at all, or (2) the very first inter-decoy interval contributes too much to the metric at the expense of all following inter-decoy intervals. We fix this problem with LogAUC by showing a simple way to choose the cutoff parameter based on the number of decoys which forces the first inter-decoy interval to always have a stable, sensible contribution to the total value. Moreover, we introduce a normalized version of LogAUC known as $\textit{enrichment score}$, which (1) enforces stability by selecting the cutoff parameter in the manner described, (2) yields scores which are more intuitively meaningful, and (3) allows reliably accurate comparison of the enrichment capacities exhibited by different ROC curves, even those produced using different numbers of decoys. Finally, we demonstrate the advantage of enrichment score over unbalanced metrics using data from a real retrospective docking study performed using the program $\textit{DOCK 3.7}$ on the target receptor TRYB1 included in the $\textit{DUDE-Z}$ benchmark.

📄 PDF Abstract BibTeX arXiv:2210.10905

Code (0)

등록된 구현이 없습니다.

Tasks

Molecular Docking

Similar Papers 제목 키워드 기반

NetScore: Towards Universal Metrics for Large-scale Performance Analysis of Deep Neural Networks for Practical On-Device Edge Usage

2018-06-14 · Alexander Wong

Much of the focus in the design of deep neural networks has been on improving accuracy, leading to more powerful yet highly complex network architectures that are difficult to deploy in practical scenarios, particularly …

image-classificationImage ClassificationObject Recognition

Multimodal Benchmarking and Recommendation of Text-to-Image Generation Models

2025-05-06 · Kapil Wanaskar, Gaytri Jena, Magdalini Eirinaki

This work presents an open-source unified benchmarking and evaluation framework for text-to-image generation models, with a particular focus on the impact of metadata augmented prompts. Leveraging the DeepFashion-MultiMo…

BenchmarkingImage GenerationMLLM Aesthetic EvaluationModel Selection+4

On Quantitative Evaluations of Counterfactuals

2021-10-30 · Frederik Hvilshøj, Alexandros Iosifidis, Ira Assent

As counterfactual examples become increasingly popular for explaining decisions of deep learning models, it is essential to understand what properties quantitative evaluation metrics do capture and equally important what…

counterfactual

A Quantitative Evaluation Framework for Missing Value Imputation Algorithms

2013-11-10 · Vinod Nair, Rahul Kidambi, Sundararajan Sellamanickam, S. Sathiya Keerthi 외

We consider the problem of quantitatively evaluating missing value imputation algorithms. Given a dataset with missing values and a choice of several imputation algorithms to fill them in, there is currently no principle…

ImputationMissing Values

VCU at Semeval-2016 Task 14: Evaluating definitional-based similarity measure for semantic taxonomy enrichment

2016-06-01 · SEMEVAL 2016 6 · Bridget McInnes
Information RetrievalSemantic Textual SimilarityWord Sense Disambiguation