paper-with-me

홈 › Papers

Similarity-Preserving Knowledge Distillation

2019-07-23 · ICCV 2019 10 · Frederick Tung, Greg Mori

Knowledge distillation is a widely applicable technique for training a student neural network under the guidance of a trained teacher network. For example, in neural network compression, a high-capacity teacher is distilled to train a compact student; in privileged learning, a teacher trained with privileged data is distilled to train a student without access to that data. The distillation loss determines how a teacher's knowledge is captured and transferred to the student. In this paper, we propose a new form of knowledge distillation loss that is inspired by the observation that semantically similar inputs tend to elicit similar activation patterns in a trained network. Similarity-preserving knowledge distillation guides the training of a student network such that input pairs that produce similar (dissimilar) activations in the teacher network produce similar (dissimilar) activations in the student network. In contrast to previous distillation methods, the student is not required to mimic the representation space of the teacher, but rather to preserve the pairwise similarities in its own representation space. Experiments on three public datasets demonstrate the potential of our approach.

📄 PDF Abstract BibTeX arXiv:1907.09682

Code (1)

yoshitomo-matsubara/torchdistill pytorch

Tasks

Knowledge DistillationNeural Network Compression

Methods 이 논문이 사용한 방법론

Knowledge Distillation A very simple way to improve the performance of almost any machine learning algorithm is to train many different models on the same data and then to average their predictions.…

Similar Papers 제목 키워드 기반

A Novel Self-Knowledge Distillation Approach with Siamese Representation Learning for Action Recognition

2022-09-03 · Duc-Quang Vu, Trang Phung, Jia-Ching Wang

Knowledge distillation is an effective transfer of knowledge from a heavy network (teacher) to a small network (student) to boost students' performance. Self-knowledge distillation, the special case of knowledge distilla…

Action RecognitionKnowledge DistillationRepresentation LearningSelf-Knowledge Distillation

On Representation Knowledge Distillation for Graph Neural Networks

2021-11-09 · Chaitanya K. Joshi, Fayao Liu, Xu Xun, Jie Lin 외

Knowledge distillation is a learning paradigm for boosting resource-efficient graph neural networks (GNNs) using more expressive yet cumbersome teacher models. Past work on distillation for GNNs proposed the Local Struct…

Contrastive LearningKnowledge Distillation

Categorical Relation-Preserving Contrastive Knowledge Distillation for Medical Image Classification

2021-07-07 · Xiaohan Xing, Yuenan Hou, Hang Li, Yixuan Yuan 외

The amount of medical images for training deep classification models is typically very scarce, making these deep models prone to overfit the training data. Studies showed that knowledge distillation (KD), especially the …

Classificationimage-classificationImage ClassificationKnowledge Distillation+2

Two-Step Knowledge Distillation for Tiny Speech Enhancement

2023-09-15 · Rayan Daod Nathoo, Mikolaj Kegler, Marko Stamenovic

Tiny, causal models are crucial for embedded audio machine learning applications. Model compression can be achieved via distilling knowledge from a large teacher into a smaller student model. In this work, we propose a n…

Knowledge DistillationModel CompressionSpeech Enhancement

D3still: Decoupled Differential Distillation for Asymmetric Image Retrieval

2024-01-01 · CVPR 2024 1 · Yi Xie, Yihong Lin, Wenjie Cai, Xuemiao Xu 외

Existing methods for asymmetric image retrieval employ a rigid pairwise similarity constraint between the query network and the larger gallery network. However these one-to-one constraint approaches often fail to mai…

Image RetrievalRetrieval