paper-with-me

Papers

ProFSA: Self-supervised Pocket Pretraining via Protein Fragment-Surroundings Alignment

2023-10-11 · Bowen Gao, Yinjun Jia, Yuanle Mo, Yuyan Ni, WeiYing Ma, ZhiMing Ma, Yanyan Lan

Pocket representations play a vital role in various biomedical applications, such as druggability estimation, ligand affinity prediction, and de novo drug design. While existing geometric features and pretrained representations have demonstrated promising results, they usually treat pockets independent of ligands, neglecting the fundamental interactions between them. However, the limited pocket-ligand complex structures available in the PDB database (less than 100 thousand non-redundant pairs) hampers large-scale pretraining endeavors for interaction modeling. To address this constraint, we propose a novel pocket pretraining approach that leverages knowledge from high-resolution atomic protein structures, assisted by highly effective pretrained small molecule representations. By segmenting protein structures into drug-like fragments and their corresponding pockets, we obtain a reasonable simulation of ligand-receptor interactions, resulting in the generation of over 5 million complexes. Subsequently, the pocket encoder is trained in a contrastive manner to align with the representation of pseudo-ligand furnished by some pretrained small molecule encoders. Our method, named ProFSA, achieves state-of-the-art performance across various tasks, including pocket druggability prediction, pocket matching, and ligand binding affinity prediction. Notably, ProFSA surpasses other pretraining methods by a substantial margin. Moreover, our work opens up a new avenue for mitigating the scarcity of protein-ligand complex data through the utilization of high-quality and diverse protein structure databases.

📄 PDF Abstract BibTeX arXiv:2310.07229

Code (1)

bowen-gao/ProFSA pytorch

Tasks

Drug Design

Methods 이 논문이 사용한 방법론

ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

CoSP: Co-supervised pretraining of pocket and ligand

2022-06-23 · Zhangyang Gao, Cheng Tan, Lirong Wu, Stan Z. Li

Can we inject the pocket-ligand interaction knowledge into the pre-trained model and jointly learn their chemical space? Pretraining molecules and proteins has attracted considerable attention in recent years, while most…

Contrastive LearningSpecificity

Uni-Mol: A Universal 3D Molecular Representation Learning Framework

2022-09-08 · ChemRxiv 2022 9 · Gengmo Zhou, Zhifeng Gao, Qiankun Ding, Hang Zheng 외

Molecular representation learning (MRL) has gained tremendous attention due to its critical role in learning from limited supervised data for applications like drug design. In most MRL methods, molecules are treated as 1…

3D geometry3D Geometry PredictionDrug DesignMolecular Property Prediction+5

Harmonic Self-Conditioned Flow Matching for Multi-Ligand Docking and Binding Site Design

2023-10-09 · Hannes Stärk, Bowen Jing, Regina Barzilay, Tommi Jaakkola

A significant amount of protein function requires binding small molecules, including enzymatic catalysis. As such, designing binding pockets for small molecules has several impactful applications ranging from drug synthe…

The Docking Game: Loop Self-Play for Fast, Dynamic, and Accurate Prediction of Flexible Protein-Ligand Binding

2025-08-07 · Youzhi Zhang, Yufei Li, Gaofeng Meng, Hongbin Liu 외 arxiv

Molecular docking is a crucial aspect of drug discovery, as it predicts the binding interactions between small-molecule ligands and protein pockets. However, current multi-task learning models for docking often show infe…

Multi-Task LearningDrug Discovery

SiteFerret: beyond simple pocket identification in proteins

2022-12-22 · Luca Gagliardi, Walter Rocchia

We present a novel method for the automatic detection of pockets on protein molecular surfaces. The algorithm is based on an ad hoc hierarchical clustering of virtual SES probe spheres obtained from the geometrical primi…

Anomaly DetectionClustering