paper-with-me

Papers

Best practices for constructing, preparing, and evaluating protein-ligand binding affinity benchmarks

2021-05-13 · David F. Hahn, Christopher I. Bayly, Hannah E. Bruce Macdonald, John D. Chodera, Vytautas Gapsys, Antonia S. J. S. Mey, David L. Mobley, Laura Perez Benito, Christina E. M. Schindler, Gary Tresadern, Gregory L. Warren

Free energy calculations are rapidly becoming indispensable in structure-enabled drug discovery programs. As new methods, force fields, and implementations are developed, assessing their expected accuracy on real-world systems (benchmarking) becomes critical to provide users with an assessment of the accuracy expected when these methods are applied within their domain of applicability, and developers with a way to assess the expected impact of new methodologies. These assessments require construction of a benchmark - a set of well-prepared, high quality systems with corresponding experimental measurements designed to ensure the resulting calculations provide a realistic assessment of expected performance when these methods are deployed within their domains of applicability. To date, the community has not yet adopted a common standardized benchmark, and existing benchmark reports suffer from a myriad of issues, including poor data quality, limited statistical power, and statistically deficient analyses, all of which can conspire to produce benchmarks that are poorly predictive of real-world performance. Here, we address these issues by presenting guidelines for (1) curating experimental data to develop meaningful benchmark sets, (2) preparing benchmark inputs according to best practices to facilitate widespread adoption, and (3) analysis of the resulting predictions to enable statistically meaningful comparisons among methods and force fields.

📄 PDF Abstract BibTeX arXiv:2105.06222

Code (2)

openforcefield/FE-Benchmarks-Best-Practices 공식 구현
openforcefield/protein-ligand-benchmark 공식 구현

Tasks

BenchmarkingDrug Discovery

Similar Papers 제목 키워드 기반

Deep learning for reconstructing protein structures from cryo-EM density maps: recent advances and future directions

2022-09-16 · Nabin Giri, Raj S. Roy, Jianlin Cheng

Cryo-Electron Microscopy (cryo-EM) has emerged as a key technology to determine the structure of proteins, particularly large protein complexes and assemblies in recent years. A key challenge in cryo-EM data analysis is …

Deep Learning

Data Quality Over Quantity: Pitfalls and Guidelines for Process Analytics

2022-11-11 · Lim C. Siang, Shams Elnawawi, Lee D. Rippon, Daniel L. O'Connor 외

A significant portion of the effort involved in advanced process control, process analytics, and machine learning involves acquiring and preparing data. Literature often emphasizes increasingly complex modelling techniqu…

Time SeriesTime Series Analysis

Automating MD simulations for Proteins using Large language Models: NAMD-Agent

2025-07-10 · Achuth Chandrasekhar, Amir Barati Farimani

Molecular dynamics simulations are an essential tool in understanding protein structure, dynamics, and function at the atomic level. However, preparing high quality input files for MD simulations can be a time consuming …

Code GenerationNavigate

Aligning Large Language Models and Geometric Deep Models for Protein Representation

2024-11-08 · Dong Shu, Bingbing Duan, Kai Guo, Kaixiong Zhou 외

Latent representation alignment has become a foundational technique for constructing multimodal large language models (MLLM) by mapping embeddings from different modalities into a shared space, often aligned with the emb…

Floating Anchor Diffusion Model for Multi-motif Scaffolding

2024-06-05 · Ke Liu, Weian Mao, Shuaike Shen, Xiaoran Jiao 외

Motif scaffolding seeks to design scaffold structures for constructing proteins with functions derived from the desired motif, which is crucial for the design of vaccines and enzymes. Previous works approach the problem …

model