paper-with-me

홈 › Papers

BERT Learns (and Teaches) Chemistry

2020-07-11 · Josh Payne, Mario Srouji, Dian Ang Yap, Vineet Kosaraju

Modern computational organic chemistry is becoming increasingly data-driven. There remain a large number of important unsolved problems in this area such as product prediction given reactants, drug discovery, and metric-optimized molecule synthesis, but efforts to solve these problems using machine learning have also increased in recent years. In this work, we propose the use of attention to study functional groups and other property-impacting molecular substructures from a data-driven perspective, using a transformer-based model (BERT) on datasets of string representations of molecules and analyzing the behavior of its attention heads. We then apply the representations of functional groups and atoms learned by the model to tackle problems of toxicity, solubility, drug-likeness, and synthesis accessibility on smaller datasets using the learned representations as features for graph convolution and attention models on the graph structure of molecules, as well as fine-tuning of BERT. Finally, we propose the use of attention visualization as a helpful tool for chemistry practitioners and students to quickly identify important substructures in various chemical properties.

📄 PDF Abstract BibTeX arXiv:2007.16012

Code (0)

등록된 구현이 없습니다.

Tasks

Drug Discovery

Methods 이 논문이 사용한 방법론

Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
Dense Connections Dense Connections, or Fully Connected Connections, are a type of layer in a deep neural network that use a linear operation where every input is connected to every output…
WordPiece 설명 없음
Residual Connection 설명 없음
Linear Warmup With Linear Decay Linear Warmup With Linear Decay is a learning rate schedule in which we increase the learning rate linearly for $n$ updates and then linearly decay afterwards.
Refunds@Expedia|||How do I get a full refund from Expedia? “How do I get a full refund from Expedia? How do I get a full refund from Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Quick Help &…
Layer Normalization Unlike batch normalization, Layer Normalization directly estimates the normalization statistics from the summed inputs…
Attention 설명 없음

Similar Papers 제목 키워드 기반

Learning by Teaching, with Application to Neural Architecture Search

2021-03-11 · Parth Sheth, Yueyu Jiang, Pengtao Xie

In human learning, an effective skill in improving learning outcomes is learning by teaching: a learner deepens his/her understanding of a topic by teaching this topic to others. In this paper, we aim to borrow this teac…

Neural Architecture Search

Transfer Learning for Scientific Data Chain Extraction in Small Chemical Corpus with BERT-CRF Model

2019-05-13 · Na Pang, Li Qian, Weimin Lyu, Jin-Dong Yang

Computational chemistry develops fast in recent years due to the rapid growth and breakthroughs in AI. Thanks for the progress in natural language processing, researchers can extract more fine-grained knowledge in public…

Computational chemistryEntity Extraction using GANNERTransfer Learning

Bi-semantic Chemical Embedder for Joint Representation Learning of SMILES and Natural Language

2026-08-04 · David Ming Segura, Jeremy Goumaz, Joshua W. Sin, Bojana Ranković 외 arxiv

Transformer models have revolutionized natural language processing (NLP), and text-based molecular representations like SMILES have successfully extended these architectures to chemistry. However, domain-adaptive pre-tra…

Molecular Property PredictionRepresentation LearningContrastive Learning

Beer-Lambert Autoencoder for Unsupervised Stain Representation Learning and Deconvolution in Multi-immunohistochemical Brightfield Histology Images

2026-01-16 · Mark Eastwood, Thomas McKee, Zedong Hu, Sabine Tejpar 외 arxiv

Separating the contributions of individual chromogenic stains in RGB histology whole slide images (WSIs) is essential for stain normalization, quantitative assessment of marker expression, and cell-level readouts in immu…

Representation Learning

Chunk Twice, Embed Once: A Systematic Study of Segmentation and Representation Trade-offs in Chemistry-Aware Retrieval-Augmented Generation

2025-06-13 · Mahmoud Amiri, Thomas Bocklitz

Retrieval-Augmented Generation (RAG) systems are increasingly vital for navigating the ever-expanding body of scientific literature, particularly in high-stakes domains such as chemistry. Despite the promise of RAG, foun…

ChunkingRAGRetrievalRetrieval-augmented Generation