paper-with-me

Papers

Reinforcement Learning with LLM-Guided Action Spaces for Synthesizable Lead Optimization

2026-04-09 · Tao Li, Kaiyuan Hou, Tuan Vinh, Monika Raj, Zhichun Guo, Carl Yang arxiv

Lead optimization in drug discovery requires improving therapeutic properties while ensuring that molecular modifications correspond to feasible synthetic routes. Existing approaches either prioritize property scores without enforcing synthesizability, or rely on expensive enumeration over large reaction networks, while direct application of Large Language Models (LLMs) to molecular generation frequently produces chemically invalid structures. We introduce MolReAct, a framework that formulates lead optimization as a Markov Decision Process over a synthesis-constrained action space defined by validated reaction templates. A tool-augmented LLM agent serves as a dynamic reaction environment, invoking specialized chemical analysis tools to identify reactive sites and functional groups and proposing a compact set of chemically grounded transformations from matched templates. A dedicated policy model trained via Group Relative Policy Optimization (GRPO) selects among these constrained actions to maximize long-term oracle reward across multi-step trajectories, with a SMILES-based caching mechanism reducing end-to-end optimization time by approximately 43%. Across 13 property optimization tasks from the Therapeutic Data Commons and one structure-based docking task, MolReAct achieves an average Top-10 score of 0.571, the highest among all baselines, ranking first or second on 13 of 14 tasks and attaining the best sample efficiency on 9 of 14 tasks. By grounding every optimization step in validated reaction templates, MolReAct produces molecules that are not only property-improved but each accompanied by an explicit template-grounded synthetic pathway.

📄 PDF Abstract BibTeX arXiv:2604.07669

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement LearningDrug Discovery

Similar Papers 제목 키워드 기반

Synthesizable Molecular Generation via Soft-constrained GFlowNets with Rich Chemical Priors

2026-02-04 · Hyeonah Kim, Minsu Kim, Celine Roget, Dionessa Biton 외 arxiv

The application of generative models for experimental drug discovery campaigns is severely limited by the difficulty of designing molecules de novo that can be synthesized in practice. Previous works have leveraged Gener…

Contrastive LearningDrug Discovery

Bridging Theory and Experiment in Materials Discovery: Machine-Learning-Assisted Prediction of Synthesizable Structures

2025-05-14 · Yu Xin, Peng Liu, Zhuohang Xie, Wenhui Mi 외

Even though thermodynamic energy-based crystal structure prediction (CSP) has revolutionized materials discovery, the energy-driven CSP approaches often struggle to identify experimentally realizable metastable materials…

System of Agentic AI for the Discovery of Metal-Organic Frameworks

2025-04-18 · Theo Jaffrelot Inizan, Sherry Yang, Aaron Kaplan, Yen-hsu Lin 외

Generative models and machine learning promise accelerated material discovery in MOFs for CO2 capture and water harvesting but face significant challenges navigating vast chemical spaces while ensuring synthetizability. …

Language ModelingLanguage ModellingLarge Language Model

Generating readily synthesizable small molecule fluorophore scaffolds with reinforcement learning

2026-01-12 · Ruhi Sayana, Kate Callon, Jennifer Xu, Jonathan Deutsch 외 arxiv

Developing new fluorophores for advanced imaging techniques requires exploring new chemical space. While generative AI approaches have shown promise in designing novel dye scaffolds, prior efforts often produced syntheti…

Reinforcement Learning

SynLlama: Generating Synthesizable Molecules and Their Analogs with Large Language Models

2025-03-16 · Kunyang Sun, Dorian Bagni, Joseph M. Cavanagh, Yingze Wang 외

Generative machine learning models for small molecule drug discovery have shown immense promise, but many molecules they generate are too difficult to synthesize, making them impractical for further investigation or deve…

Drug Discovery