paper-with-me

Papers

Unveiling Latent Knowledge in Chemistry Language Models through Sparse Autoencoders

2025-12-08 · Jaron Cohen, Alexander G. Hasson, Sara Tanovic arxiv

Since the advent of machine learning, interpretability has remained a persistent challenge, becoming increasingly urgent as generative models support high-stakes applications in drug and material discovery. Recent advances in large language model (LLM) architectures have yielded chemistry language models (CLMs) with impressive capabilities in molecular property prediction and molecular generation. However, how these models internally represent chemical knowledge remains poorly understood. In this work, we extend sparse autoencoder techniques to uncover and examine interpretable features within CLMs. Applying our methodology to the Foundation Models for Materials (FM4M) SMI-TED chemistry foundation model, we extract semantically meaningful latent features and analyse their activation patterns across diverse molecular datasets. Our findings reveal that these models encode a rich landscape of chemical concepts. We identify correlations between specific latent features and distinct domains of chemical knowledge, including structural motifs, physicochemical properties, and pharmacological drug classes. Our approach provides a generalisable framework for uncovering latent knowledge in chemistry-focused AI systems. This work has implications for both foundational understanding and practical deployment; with the potential to accelerate computational chemistry research.

📄 PDF Abstract BibTeX arXiv:2512.08077

Code (0)

등록된 구현이 없습니다.

Tasks

Molecular Property Prediction

Similar Papers 제목 키워드 기반

Latent Anomaly Knowledge Excavation: Unveiling Sparse Sensitive Neurons in Vision-Language Models

2026-04-09 · Shaotian Li, Shangze Li, Chuancheng Shi, Wenhua Wu 외 arxiv

Large-scale vision-language models (VLMs) exhibit remarkable zero-shot capabilities, yet the internal mechanisms driving their anomaly detection (AD) performance remain poorly understood. Current methods predominantly tr…

Anomaly Detection

From Generalist to Specialist: A Survey of Large Language Models for Chemistry

2024-12-28 · Yang Han, Ziping Wan, Lu Chen, Kai Yu 외

Large Language Models (LLMs) have significantly transformed our daily life and established a new paradigm in natural language processing (NLP). However, the predominant pretraining of LLMs on extensive web-based texts re…

scientific discoverySurvey

From Words to Molecules: A Survey of Large Language Models in Chemistry

2024-02-02 · Chang Liao, Yemin Yu, Yu Mei, Ying WEI

In recent years, Large Language Models (LLMs) have achieved significant success in natural language processing (NLP) and various interdisciplinary areas. However, applying LLMs to chemistry is a complex task that require…

Continual Learning

Interpretable and Explainable Machine Learning for Materials Science and Chemistry

2021-11-01 · Felipe Oviedo, Juan Lavista Ferres, Tonio Buonassisi, Keith Butler

While the uptake of data-driven approaches for materials science and chemistry is at an exciting, early stage, to realise the true potential of machine learning models for successful scientific discovery, they must have …

BIG-bench Machine LearningInterpretable Machine Learningscientific discovery

MOOSE-Chem: Large Language Models for Rediscovering Unseen Chemistry Scientific Hypotheses

2024-10-09 · Zonglin Yang, Wanhao Liu, Ben Gao, Tong Xie 외

Scientific discovery plays a pivotal role in advancing human society, and recent progress in large language models (LLMs) suggests their potential to accelerate this process. However, it remains unclear whether LLMs can …

scientific discoveryvalid