paper-with-me

홈 › Papers

Faster and more diverse de novo molecular optimization with double-loop reinforcement learning using augmented SMILES

2022-10-22 · Esben Jannik Bjerrum, Christian Margreitter, Thomas Blaschke, Raquel López-Ríos de Castro

Using generative deep learning models and reinforcement learning together can effectively generate new molecules with desired properties. By employing a multi-objective scoring function, thousands of high-scoring molecules can be generated, making this approach useful for drug discovery and material science. However, the application of these methods can be hindered by computationally expensive or time-consuming scoring procedures, particularly when a large number of function calls are required as feedback in the reinforcement learning optimization. Here, we propose the use of double-loop reinforcement learning with simplified molecular line entry system (SMILES) augmentation to improve the efficiency and speed of the optimization. By adding an inner loop that augments the generated SMILES strings to non-canonical SMILES for use in additional reinforcement learning rounds, we can both reuse the scoring calculations on the molecular level, thereby speeding up the learning process, as well as offer additional protection against mode collapse. We find that employing between 5 and 10 augmentation repetitions is optimal for the scoring functions tested and is further associated with an increased diversity in the generated compounds, improved reproducibility of the sampling runs and the generation of molecules of higher similarity to known ligands.

📄 PDF Abstract BibTeX arXiv:2210.12458

Code (0)

등록된 구현이 없습니다.

Tasks

DiversityDrug Discoveryreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Methods 이 논문이 사용한 방법론

SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…

Similar Papers 제목 키워드 기반

MolRGen: A Training and Evaluation Setting for De Novo Molecular Generation with Reasonning Models

2026-03-18 · Philippe Formont, Maxime Darrin, Ismail Ben Ayed, Pablo Piantanida arxiv

Recent reasoning-based large language models have shown strong performance on tasks with verifiable outcomes, but their use in de novo molecular generation remains limited by the lack of training environments where rewar…

Reinforcement Learning

Deep Lead Optimization: Leveraging Generative AI for Structural Modification

2024-04-30 · Odin Zhang, Haitao Lin, HUI ZHANG, Huifeng Zhao 외

The idea of using deep-learning-based molecular generation to accelerate discovery of drug candidates has attracted extraordinary attention, and many deep generative models have been developed for automated drug design, …

Drug Design

NovoMolGen: Rethinking Molecular Language Model Pretraining

2025-08-19 · Kamran Chitsaz, Roshan Balaji, Quentin Fournier, Nirav Pravinbhai Bhatt 외 arxiv

Designing de-novo molecules with desired property profiles requires efficient exploration of the vast chemical space ranging from $10^{23}$ to $10^{60}$ possible synthesizable candidates. While various deep generative mo…

Generative Multi-Objective Bayesian Optimization with Scalable Batch Evaluations for Sample-Efficient De Novo Molecular Design

2025-12-19 · Madhav R. Muthyala, Farshud Sorourifar, Tianhong Tan, You Peng 외 arxiv

Designing molecules that must satisfy multiple, often conflicting objectives is a central challenge in molecular discovery. The enormous size of chemical space and the cost of high-fidelity simulations have driven the de…

Diverse Mini-Batch Selection in Reinforcement Learning for Efficient Chemical Exploration in de novo Drug Design

2025-06-26 · Hampus Gummesson Svensson, Ola Engkvist, Jon Paul Janet, Christian Tyrchan 외

In many real-world applications, evaluating the goodness of instances is often costly and time-consuming, e.g., human feedback and physics simulations, in contrast to proposing new instances. In particular, this is even …

Drug DesignDrug DiscoveryPoint Processes