paper-with-me

Papers

Tanimoto Random Features for Scalable Molecular Machine Learning

2023-06-26 · NeurIPS 2023 11 · Austin Tripp, Sergio Bacallado, Sukriti Singh, José Miguel Hernández-Lobato

The Tanimoto coefficient is commonly used to measure the similarity between molecules represented as discrete fingerprints, either as a distance metric or a positive definite kernel. While many kernel methods can be accelerated using random feature approximations, at present there is a lack of such approximations for the Tanimoto kernel. In this paper we propose two kinds of novel random features to allow this kernel to scale to large datasets, and in the process discover a novel extension of the kernel to real-valued vectors. We theoretically characterize these random features, and provide error bounds on the spectral norm of the Gram matrix. Experimentally, we show that these random features are effective at approximating the Tanimoto coefficient of real-world datasets and are useful for molecular property prediction and optimization tasks.

📄 PDF Abstract BibTeX arXiv:2306.14809

Code (1)

austint/tanimoto-random-features-neurips23 공식 구현

Tasks

Molecular Property PredictionProperty Prediction

Similar Papers 제목 키워드 기반

QBioFusion-QSAR: Morgan-Anchored Quantum Multiple Kernel Learning for Small-Data Ligand Classification

2026-06-19 · Azadeh Alavi, Fatemeh Kouchmeshki, Muhammad Usman, Jessica Holien arxiv

Small quantitative structure-activity relationship (QSAR) studies are difficult when close molecular analogues have different activity labels. This paper asks whether a quantum kernel can add similarity information to a …

Weighted Tanimoto Coefficient for 3D Molecule Structure Similarity Measurement

2018-06-10 · Siti Asmah Bero, Azah Kamilah Muda, Yun-Huoy Choo, Noor Azilah Muda 외

Similarity searching of molecular structure has been an important application in the Chemoinformatics, especially in drug discovery. Similarity searching is a common method used for identification of molecular structure.…

Drug Discovery

Advancing Ligand-based Virtual Screening and Molecular Generation with Pretrained Molecular Embedding Distance

2026-04-27 · Shiyun Wa, Yifei Wang, Simone Sciabola, Ye Wang arxiv

Molecular similarity plays a central role in ligand-based drug discovery, such as virtual screening, analog searching, and goal-directed molecular generation. However, traditional similarity measures, ranging from finger…

Drug Discovery

ChemoVerse: Manifold traversal of latent spaces for novel molecule discovery

2020-09-29 · Harshdeep Singh, Nicholas McCarthy, Qurrat Ul Ain, Jeremiah Hayes

In order to design a more potent and effective chemical entity, it is essential to identify molecular structures with the desired chemical properties. Recent advances in generative models using neural networks and machin…

Heuristic Search

When Does Context Help? A Systematic Study of Target-Conditional Molecular Property Prediction

2026-04-08 · Bryan Cheng, Jasper Zhang arxiv

We present the first systematic study of when target context helps molecular property prediction, evaluating context conditioning across 10 diverse protein families, 4 fusion architectures, data regimes spanning 67-9,409…

Molecular Property Prediction