Meta-Learning Initializations for Low-Resource Drug Discovery
Building in silico models to predict chemical properties and activities is a crucial step in drug discovery. However, drug discovery projects are often characterized by limited labeled data, hindering the applications of deep learning in this setting. Meanwhile advances in meta-learning have enabled state-of-the-art performances in few-shot learning benchmarks, naturally prompting the question: Can meta-learning improve deep learning performance in low-resource drug discovery projects? In this work, we assess the efficiency of the Model-Agnostic Meta-Learning (MAML) algorithm - along with its variants FO-MAML and ANIL - at learning to predict chemical properties and activities. Using the ChEMBL20 dataset to emulate low-resource settings, our benchmark shows that meta-initializations perform comparably to or outperform multi-task pre-training baselines on 16 out of 20 in-distribution tasks and on all out-of-distribution tasks, providing an average improvement in AUPRC of 7.2% and 14.9% respectively. Finally, we observe that meta-initializations consistently result in the best performing models across fine-tuning sets with $k \in \{16, 32, 64, 128, 256\}$ instances.
Code (1)
Tasks
Drug DiscoveryFew-Shot LearningMeta-LearningSimilar Papers 제목 키워드 기반
Meta-Learning GNN Initializations for Low-Resource Molecular Property Prediction
Building $\textit{in silico}$ models to predict chemical properties and activities is a crucial step in drug discovery. However, limited labeled data often hinders the application of deep learning in this setting. Meanwh…
Drug DiscoveryFew-Shot LearningMeta-LearningMolecular Property Prediction+1Artificial Intelligence in Drug Discovery: Applications and Techniques
Artificial intelligence (AI) has been transforming the practice of drug discovery in the past decade. Various AI techniques have been used in a wide range of applications, such as virtual screening and drug design. In th…
Drug DesignDrug DiscoveryMolecular Property PredictionProperty Prediction+1COMO: A Pipeline for Multi-Omics Data Integration in Metabolic Modeling and Drug Discovery
Identifying potential drug targets using metabolic modeling requires integrating multiple modeling methods and heterogenous biological datasets, which can be challenging without sophisticated tools. We developed COMO, a …
Data IntegrationDrug DiscoverySMILES-Mamba: Chemical Mamba Foundation Models for Drug ADMET Prediction
In drug discovery, predicting the absorption, distribution, metabolism, excretion, and toxicity (ADMET) properties of small-molecule drugs is critical for ensuring safety and efficacy. However, the process of accurately …
Drug DiscoveryMambaMolecular Property PredictionProperty Prediction+1The Clinical Trials Puzzle: How Network Effects Limit Drug Discovery
The depth of knowledge offered by post-genomic medicine has carried the promise of new drugs, and cures for multiple diseases. To explore the degree to which this capability has materialized, we extract meta-data from 35…
Drug Discovery