$\text{M}^{2}$LLM: Multi-view Molecular Representation Learning with Large Language Models
Accurate molecular property prediction is a critical challenge with wide-ranging applications in chemistry, materials science, and drug discovery. Molecular representation methods, including fingerprints and graph neural networks (GNNs), achieve state-of-the-art results by effectively deriving features from molecular structures. However, these methods often overlook decades of accumulated semantic and contextual knowledge. Recent advancements in large language models (LLMs) demonstrate remarkable reasoning abilities and prior knowledge across scientific domains, leading us to hypothesize that LLMs can generate rich molecular representations when guided to reason in multiple perspectives. To address these gaps, we propose $\text{M}^{2}$LLM, a multi-view framework that integrates three perspectives: the molecular structure view, the molecular task view, and the molecular rules view. These views are fused dynamically to adapt to task requirements, and experiments demonstrate that $\text{M}^{2}$LLM achieves state-of-the-art performance on multiple benchmarks across classification and regression tasks. Moreover, we demonstrate that representation derived from LLM achieves exceptional performance by leveraging two core functionalities: the generation of molecular embeddings through their encoding capabilities and the curation of molecular features through advanced reasoning processes.
Code (0)
등록된 구현이 없습니다.
Tasks
Molecular Property PredictionRepresentation LearningDrug DiscoverySimilar Papers 제목 키워드 기반
MV-CLAM: Multi-View Molecular Interpretation with Cross-Modal Projection via Language Model
Human expertise in chemistry and biomedicine relies on contextual molecular understanding, a capability that large language models (LLMs) can extend through fine-grained alignment between molecular structures and text. R…
cross-modal alignmentLanguage ModelingLanguage ModellingLearning Multi-view Molecular Representations with Structured and Unstructured Knowledge
Capturing molecular knowledge with representation learning approaches holds significant potential in vast scientific fields such as chemistry and life science. An effective and generalizable molecular representation is e…
Knowledge GraphsMolecular Property Predictionmolecular representationProperty Prediction+1Multi-view biomedical foundation models for molecule-target and property prediction
Foundation models applied to bio-molecular space hold promise to accelerate drug discovery. Molecular representation is key to building such models. Previous works have typically focused on a single representation or vie…
Drug Discoverymolecular representationProperty PredictionCROP: Integrating Topological and Spatial Structures via Cross-View Prefixes for Molecular LLMs
Recent advances in molecular science have been propelled significantly by large language models (LLMs). However, their effectiveness is limited when relying solely on molecular sequences, which fail to capture the comple…
Molecule CaptioningMolecular Property Prediction by Semantic-invariant Contrastive Learning
Contrastive learning have been widely used as pretext tasks for self-supervised pre-trained molecular representation learning models in AI-aided drug design and discovery. However, exiting methods that generate molecular…
Contrastive LearningDrug DesignMolecular Property Predictionmolecular representation+3