Why Deep Models Often cannot Beat Non-deep Counterparts on Molecular Property Prediction?
Molecular property prediction (MPP) is a crucial task in the drug discovery pipeline, which has recently gained considerable attention thanks to advances in deep neural networks. However, recent research has revealed that deep models struggle to beat traditional non-deep ones on MPP. In this study, we benchmark 12 representative models (3 non-deep models and 9 deep models) on 14 molecule datasets. Through the most comprehensive study to date, we make the following key observations: \textbf{(\romannumeral 1)} Deep models are generally unable to outperform non-deep ones; \textbf{(\romannumeral 2)} The failure of deep models on MPP cannot be solely attributed to the small size of molecular datasets. What matters is the irregular molecule data pattern; \textbf{(\romannumeral 3)} In particular, tree models using molecular fingerprints as inputs tend to perform better than other competitors. Furthermore, we conduct extensive empirical investigations into the unique patterns of molecule data and inductive biases of various models underlying these phenomena.
Code (0)
등록된 구현이 없습니다.
Tasks
Drug DiscoveryMolecular Property PredictionProperty PredictionSimilar Papers 제목 키워드 기반
Understanding the Limitations of Deep Models for Molecular property prediction: Insights and Solutions
Molecular Property Prediction (MPP) is a crucial task in the AI-driven Drug Discovery (AIDD) pipeline, which has recently gained considerable attention thanks to advancements in deep learning. However, recent research ha…
Pre-training Transformers for Molecular Property Prediction Using Reaction Prediction
Molecular property prediction is essential in chemistry, especially for drug discovery applications. However, available molecular property data is often limited, encouraging the transfer of information from related data.…
Drug DiscoveryMolecular Property Predictionmolecular representationPrediction+3Atomic and Subgraph-aware Bilateral Aggregation for Molecular Representation Learning
Molecular representation learning is a crucial task in predicting molecular properties. Molecules are often modeled as graphs where atoms and chemical bonds are represented as nodes and edges, respectively, and Graph Neu…
Molecular Property Predictionmolecular representationProperty PredictionRepresentation Learning+1When Molecular Similarity Works: Property Cliffs Reveal Hidden Errors
Accurate prediction of molecular properties underpins drug discovery and material design, yet even state-of-the-art models remain vulnerable to localized failure modes that aggregate metrics cannot detect. The places whe…
Drug DiscoveryImplicit Geometry and Interaction Embeddings Improve Few-Shot Molecular Property Prediction
Few-shot learning is a promising approach to molecular property prediction as supervised data is often very limited. However, many important molecular properties depend on complex molecular characteristics -- such as the…
Few-Shot LearningMolecular DockingMolecular Property PredictionMulti-Task Learning+3