paper-with-me

홈 › Papers

Is Self-Supervised Pretraining Good for Extrapolation in Molecular Property Prediction?

2023-08-16 · Shun Takashige, Masatoshi Hanai, Toyotaro Suzumura, LiMin Wang, Kenjiro Taura

The prediction of material properties plays a crucial role in the development and discovery of materials in diverse applications, such as batteries, semiconductors, catalysts, and pharmaceuticals. Recently, there has been a growing interest in employing data-driven approaches by using machine learning technologies, in combination with conventional theoretical calculations. In material science, the prediction of unobserved values, commonly referred to as extrapolation, is particularly critical for property prediction as it enables researchers to gain insight into materials beyond the limits of available data. However, even with the recent advancements in powerful machine learning models, accurate extrapolation is still widely recognized as a significantly challenging problem. On the other hand, self-supervised pretraining is a machine learning technique where a model is first trained on unlabeled data using relatively simple pretext tasks before being trained on labeled data for target tasks. As self-supervised pretraining can effectively utilize material data without observed property values, it has the potential to improve the model's extrapolation ability. In this paper, we clarify how such self-supervised pretraining can enhance extrapolation performance.We propose an experimental framework for the demonstration and empirically reveal that while models were unable to accurately extrapolate absolute property values, self-supervised pretraining enables them to learn relative tendencies of unobserved property values and improve extrapolation performance.

📄 PDF Abstract BibTeX arXiv:2308.08129

Code (0)

등록된 구현이 없습니다.

Tasks

Molecular Property PredictionProperty Prediction

Similar Papers 제목 키워드 기반

3D Denoisers are Good 2D Teachers: Molecular Pretraining via Denoising and Cross-Modal Distillation

2023-09-08 · Sungjun Cho, Dae-Woong Jeong, Sung Moon Ko, Jinwoo Kim 외

Pretraining molecular representations from large unlabeled data is essential for molecular property prediction due to the high cost of obtaining ground-truth labels. While there exist various 2D graph-based molecular pre…

DenoisingKnowledge DistillationMolecular Property Predictionmolecular representation+2

Learning the Neighborhood: Contrast-Free Multimodal Self-Supervised Molecular Graph Pretraining

2025-09-26 · Boshra Ariguib, Mathias Niepert, Andrei Manolache arxiv

High-quality molecular representations are essential for property prediction and molecular design, yet large labeled datasets remain scarce. While self-supervised pretraining on molecular graphs has shown promise, many e…

Representation LearningGraph Neural Network

Does GNN Pretraining Help Molecular Representation?

2022-07-13 · Ruoxi Sun, Hanjun Dai, Adams Wei Yu

Extracting informative representations of molecules using Graph neural networks (GNNs) is crucial in AI-driven drug discovery. Recently, the graph research community has been trying to replicate the success of self-super…

Drug Discoverymolecular representation

ChemBERTa-2: Towards Chemical Foundation Models

2022-09-05 · Walid Ahmad, Elana Simon, Seyone Chithrananda, Gabriel Grand 외

Large pretrained models such as GPT-3 have had tremendous impact on modern natural language processing by leveraging self-supervised learning to learn salient representations that can be used to readily finetune on a wid…

Molecular Property PredictionSelf-Supervised Learning

Multi-level Self-supervised Pretraining on Compositional Hierarchical Graph for Molecular Property Prediction

2026-05-15 · Xiayu Liu, Zhengyi Lu, Hou-biao Li arxiv

Self-supervised pretraining on molecular graphs has emerged as a promising approach for molecular property prediction, yet most existing methods operate at a single structural granularity and treat bond information as au…

Molecular Property Prediction