paper-with-me

홈 › Papers

Accurate and efficient protein embedding using multi-teacher distillation learning

2024-05-20 · Jiayu Shang, Cheng Peng, Yongxin Ji, Jiaojiao Guan, Dehan Cai, Xubo Tang, Yanni Sun

Motivation: Protein embedding, which represents proteins as numerical vectors, is a crucial step in various learning-based protein annotation/classification problems, including gene ontology prediction, protein-protein interaction prediction, and protein structure prediction. However, existing protein embedding methods are often computationally expensive due to their large number of parameters, which can reach millions or even billions. The growing availability of large-scale protein datasets and the need for efficient analysis tools have created a pressing demand for efficient protein embedding methods. Results: We propose a novel protein embedding approach based on multi-teacher distillation learning, which leverages the knowledge of multiple pre-trained protein embedding models to learn a compact and informative representation of proteins. Our method achieves comparable performance to state-of-the-art methods while significantly reducing computational costs and resource requirements. Specifically, our approach reduces computational time by ~70\% and maintains almost the same accuracy as the original large models. This makes our method well-suited for large-scale protein analysis and enables the bioinformatics community to perform protein embedding tasks more efficiently.

📄 PDF Abstract BibTeX arXiv:2405.11735

Code (1)

KennthShang/MTDP 공식 구현 pytorch

Tasks

PredictionProtein AnnotationProtein Structure Prediction

Methods 이 논문이 사용한 방법론

Ontology 설명 없음

Similar Papers 제목 키워드 기반

Investigating Knowledge Distillation Through Neural Networks for Protein Binding Affinity Prediction

2026-01-07 · Wajid Arshad Abbasi, Syed Ali Abbas, Maryum Bibi, Saiqa Andleeb 외 arxiv

The trade-off between predictive accuracy and data availability makes it difficult to predict protein--protein binding affinity accurately. The lack of experimentally resolved protein structures limits the performance of…

Knowledge Distillation

Distilled Protein Backbone Generation

2025-10-03 · Liyang Xie, Haoran Zhang, Zhendong Wang, Wesley Tansey 외 arxiv

Diffusion- and flow-based generative models have recently demonstrated strong performance in protein backbone generation tasks, offering unprecedented capabilities for de novo protein design. However, while achieving not…

Protein Design

CLIP-RD: Relative Distillation for Efficient CLIP Knowledge Distillation

2026-03-26 · Jeannie Chung, Hanna Jang, Ingyeong Yang, Uiwon Hwang 외 arxiv

CLIP aligns image and text embeddings via contrastive learning and demonstrates strong zero-shot generalization. Its large-scale architecture requires substantial computational and memory resources, motivating the distil…

Zero-shot GeneralizationKnowledge DistillationContrastive Learning

CLIP-Embed-KD: Computationally Efficient Knowledge Distillation Using Embeddings as Teachers

2024-04-09 · Lakshmi Nair

Contrastive Language-Image Pre-training (CLIP) has been shown to improve zero-shot generalization capabilities of language and vision models. In this paper, we extend CLIP for efficient knowledge distillation, by utilizi…

Knowledge DistillationZero-shot Generalization

Multi-Teacher Knowledge Distillation via Teacher-Informed Mixture Priors

2026-05-27 · Luyang Fang, Yongkai Chen, Jiazhang Cai, Ping Ma 외 arxiv

Knowledge distillation is a powerful method for model compression, enabling the efficient deployment of complex deep learning models (teachers), including large language models. However, its underlying statistical mechan…

Knowledge DistillationImage ClassificationBayesian InferenceModel Compression