paper-with-me

홈 › Papers

Extracting Molecular Properties from Natural Language with Multimodal Contrastive Learning

2023-07-22 · Romain Lacombe, Andrew Gaut, Jeff He, David Lüdeke, Kateryna Pistunova

Deep learning in computational biochemistry has traditionally focused on molecular graphs neural representations; however, recent advances in language models highlight how much scientific knowledge is encoded in text. To bridge these two modalities, we investigate how molecular property information can be transferred from natural language to graph representations. We study property prediction performance gains after using contrastive learning to align neural graph representations with representations of textual descriptions of their characteristics. We implement neural relevance scoring strategies to improve text retrieval, introduce a novel chemically-valid molecular graph augmentation strategy inspired by organic reactions, and demonstrate improved performance on downstream MoleculeNet property classification tasks. We achieve a +4.26% AUROC gain versus models pre-trained on the graph modality alone, and a +1.54% gain compared to recently proposed molecular graph/text contrastively trained MoMu model (Su et al. 2022).

📄 PDF Abstract BibTeX arXiv:2307.12996

Code (0)

등록된 구현이 없습니다.

Tasks

Contrastive LearningProperty PredictionRetrievalText Retrievalvalid

Methods 이 논문이 사용한 방법론

Contrastive Learning 설명 없음
ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

XMolCap: Advancing Molecular Captioning through Multimodal Fusion and Explainable Graph Neural Networks

2025-05-23 · IEEE Journal of Biomedical and Health Informatics 2025 5 · Duong Thanh Tran, Nguyen Doan Hieu Nguyen, Nhat Truong Pham, Rajan Rakkiyappan 외

Large language models (LLMs) have significantly advanced computational biology by enabling the integration of molecular, protein, and natural language data to accelerate drug discovery. However, existing molecular captio…

Drug DiscoveryMolecule Captioning

Instruction Multi-Constraint Molecular Generation Using a Teacher-Student Large Language Model

2024-03-20 · Peng Zhou, Jianmin Wang, Chunyan Li, Zixu Wang 외

While various models and computational tools have been proposed for structure and property analysis of molecules, generating molecules that conform to all desired structures and properties remains a challenge. Here, we i…

Drug DiscoveryKnowledge DistillationLanguage ModelingLanguage Modelling+1

Integrating Chemical Language and Molecular Graph in Multimodal Fused Deep Learning for Drug Property Prediction

2023-12-29 · Xiaohua Lu, Liangxu Xie, Lei Xu, Rongzhi Mao 외

Accurately predicting molecular properties is a challenging but essential task in drug discovery. Recently, many mono-modal deep learning methods have been successfully applied to molecular property prediction. However, …

Deep LearningDrug DiscoveryMolecular Property Predictionmolecular representation+2

LLM and GNN are Complementary: Distilling LLM for Multimodal Graph Learning

2024-06-03 · Junjie Xu, Zongyu Wu, Minhua Lin, Xiang Zhang 외

Recent progress in Graph Neural Networks (GNNs) has greatly enhanced the ability to model complex molecular structures for predicting properties. Nevertheless, molecular data encompasses more than just graph structures, …

Graph LearningLanguage ModelingLanguage ModellingLarge Language Model

Property Enhanced Instruction Tuning for Multi-task Molecule Generation with Large Language Models

2024-12-24 · Xuan Lin, Long Chen, Yile Wang, Xiangxiang Zeng 외

Large language models (LLMs) are widely applied in various natural language processing tasks such as question answering and machine translation. However, due to the lack of labeled data and the difficulty of manual annot…

Machine TranslationMolecular Property PredictionMolecule CaptioningProperty Prediction+1