paper-with-me

홈 › Papers

Multimodal Transformer-based Model for Buchwald-Hartwig and Suzuki-Miyaura Reaction Yield Prediction

2022-04-27 · Shimaa Baraka, Ahmed M. El Kerdawy

Predicting the yield percentage of a chemical reaction is useful in many aspects such as reducing wet-lab experimentation by giving the priority to the reactions with a high predicted yield. In this work we investigated the use of multiple type inputs to predict chemical reaction yield. We used simplified molecular-input line-entry system (SMILES) as well as calculated chemical descriptors as model inputs. The model consists of a pre-trained bidirectional transformer-based encoder (BERT) and a multi-layer perceptron (MLP) with a regression head to predict the yield. We experimented on two high throughput experimentation (HTE) datasets for Buchwald-Hartwig and Suzuki-Miyaura reactions. The experiments show improvements in the prediction on both datasets compared to systems using only SMILES or chemical descriptors as input. We also tested the model's performance on out-of-sample dataset splits of Buchwald-Hartwig and achieved comparable results with the state-of-the-art. In addition to predicting the yield, we demonstrated the model's ability to suggest the optimum (highest yield) reaction conditions. The model was able to suggest conditions that achieves 94% of the optimum reported yields. This proves the model to be useful in achieving the best results in the wet lab without expensive experimentation.

📄 PDF Abstract BibTeX arXiv:2204.14062

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

RxnCLF: Contrastive Transformation-Aware Reaction Foundation Model for Improved Reactivity Prediction

2026-08-06 · Yiting Zheng, Cheng Fang, Anthony Donofrio, Haote Li arxiv

Reaction yield prediction remains challenging because labeled data are scarce and reaction space is both combinatorially large and sparsely populated, limiting the generalization of existing reaction representations. Str…

Representation LearningContrastive Learning

Interpretable epistemic uncertainty decomposition in sequential generative models via polynomial chaos surrogates

2025-10-24 · Ramón Nartallo-Kaluarachchi, Shashanka Ubaru, Małgorzata J Zimoń, Dongsung Huh 외 arxiv

Sequential generative models conditioned on uncertain rewards are central to AI-driven scientific discovery, yet the epistemic uncertainty they inherit from imperfect reward estimates remains unquantified. We propagate t…

GOLLuM: Gaussian Process Optimized LLMs -- Reframing LLM Finetuning through Bayesian Optimization

2025-04-08 · Bojana Ranković, Philippe Schwaller

Large Language Models (LLMs) can encode complex relationships in their latent spaces, yet harnessing them for optimization under uncertainty remains challenging. We address this gap with a novel architecture that reframe…

Bayesian OptimizationContrastive LearningDecoder

RetroMPA: A Molecular Property-Aware Auxiliary Framework for Enhancing Retrosynthesis Prediction

2026-08-17 · Mianzhi Liu, Fan Xiao, Zhiliang Yu, Huayang Huang 외 arxiv

Retrosynthesis is a cornerstone of drug discovery and organic synthesis. While data-driven deep learning models have shown remarkable progress, they autonomously learn reaction patterns from extensive datasets with limit…

Drug Discovery

Gryffin: An algorithm for Bayesian optimization of categorical variables informed by expert knowledge

2020-03-26 · Florian Häse, Matteo Aldeghi, Riley J. Hickman, Loïc M. Roch 외

Designing functional molecules and advanced materials requires complex design choices: tuning continuous process parameters such as temperatures or flow rates, while simultaneously selecting catalysts or solvents. To dat…

Bayesian OptimizationDensity Estimation