paper-with-me

Papers

Text-Guided Multi-Property Molecular Optimization with a Diffusion Language Model

2024-10-17 · Yida Xiong, Kun Li, Weiwei Liu, Jia Wu, Bo Du, Shirui Pan, Wenbin Hu

Molecular optimization (MO) is a crucial stage in drug discovery in which task-oriented generated molecules are optimized to meet practical industrial requirements. Existing mainstream MO approaches primarily utilize external property predictors to guide iterative property optimization. However, learning all molecular samples in the vast chemical space is unrealistic for predictors. As a result, errors and noise are inevitably introduced during property prediction due to the nature of approximation. This leads to discrepancy accumulation, generalization reduction and suboptimal molecular candidates. In this paper, we propose a text-guided multi-property molecular optimization method utilizing transformer-based diffusion language model (TransDLM). TransDLM leverages standardized chemical nomenclature as semantic representations of molecules and implicitly embeds property requirements into textual descriptions, thereby preventing error propagation during diffusion process. Guided by physically and chemically detailed textual descriptions, TransDLM samples and optimizes encoded source molecules, retaining core scaffolds of source molecules and ensuring structural similarities. Moreover, TransDLM enables simultaneous sampling of multiple molecules, making it ideal for scalable, efficient large-scale optimization through distributed computation on web platforms. Furthermore, our approach surpasses state-of-the-art methods in optimizing molecular structural similarity and enhancing chemical properties on the benchmark dataset. The code is available at: https://anonymous.4open.science/r/TransDLM-A901.

📄 PDF Abstract BibTeX arXiv:2410.13597

Code (0)

등록된 구현이 없습니다.

Tasks

Drug DiscoveryLanguage ModelingLanguage ModellingProperty Prediction

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Property-Guided Molecular Generation and Optimization via Latent Flows

2026-03-27 · Alexander Arjun Lobo, Urvi Awasthi, Leonid Zhukov arxiv

Molecular discovery is increasingly framed as an inverse design problem: identifying molecular structures that satisfy desired property profiles under feasibility constraints. While recent generative models provide conti…

Uncertainty-Aware Multi-Objective Reinforcement Learning-Guided Diffusion Models for 3D De Novo Molecular Design

2025-10-24 · Lianghong Chen, Dongkyu Eugene Kim, Mike Domaratzki, Pingzhao Hu arxiv

Designing de novo 3D molecules with desirable properties remains a fundamental challenge in drug discovery and molecular engineering. While diffusion models have demonstrated remarkable capabilities in generating high-qu…

Reinforcement LearningDrug Discovery

From Single-Step Edit Response to Multi-Step Molecular Optimization

2026-05-11 · Haojie Rao, Kun Li, Yida Xiong, Jiameng Chen 외 arxiv

Conditional molecular optimization aims to edit a molecule to realize a specified property shift. In practice, structurally similar molecule data is scarce, while decisions are inherently action-level: at each step, the …

Scaffold-Conditioned Preference Triplets for Controllable Molecular Optimization with Large Language Models

2026-04-14 · Yi Xiong, Liang Xiong, Xiaohong Ji, Sen Yang 외 arxiv

Molecular property optimization is central to drug discovery, yet many deep learning methods rely on black-box scoring and offer limited control over scaffold preservation, often producing unstable or biologically implau…

Drug Discovery

POLO: Preference-Guided Multi-Turn Reinforcement Learning for Lead Optimization

2025-09-26 · Ziqing Wang, Yibo Wen, William Pattie, Xiao Luo 외 arxiv

Lead optimization in drug discovery requires efficiently navigating vast chemical space through iterative cycles to enhance molecular properties while preserving structural similarity to the original lead compound. Despi…

Reinforcement LearningInstruction FollowingDrug Discovery