PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
February 9, 20240 citationsOpen Access

Embedding Compression for Teacher-to-Student Knowledge Transfer

View Full Paper
YDYiwei DingALAlexander Lerch

Key Points

Key points are not available for this paper at this time.

Abstract

Common knowledge distillation methods require the teacher model and the student model to be trained on the same task. However, the usage of embeddings as teachers has also been proposed for different source tasks and target tasks. Prior work that uses embeddings as teachers ignores the fact that the teacher embeddings are likely to contain irrelevant knowledge for the target task. To address this problem, we propose to use an embedding compression module with a trainable teacher transformation to obtain a compact teacher embedding. Results show that adding the embedding compression module improves the classification performance, especially for unsupervised teacher embeddings. Moreover, student models trained with the guidance of embeddings show stronger generalizability.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Ding et al. (2024) studied this question.

synapsesocial.com/papers/68e7b285b6db64358770d526https://doi.org/10.48550/arxiv.2402.06761
Ask AI
Helpful
Bookmark
Share
View Full Paper