← Search

Yifeng Cai

2 accepted papers

2025

Membership and Memorization in LLM Knowledge Distillation

EMNLP 2025

Recent advances in Knowledge Distillation (KD) aim to mitigate the high computational demands of Large Language Models (LLMs) by transferring knowledge from a large ”teacher” to a smaller ”student” model. However, students may inherit the teacher’s privacy when the teacher is trained on private data

Cited by 0SourcePDFScholar
2021

TransTailor: Pruning the Pre-trained Model for Improved Transfer Learning

AAAI 2021technical

The increasing of pre-trained models has significantly facilitated the performance on limited data tasks with transfer learning. However, progress on transfer learning mainly focuses on optimizing the weights of pre-trained models, which ignores the structure mismatch between the model and the targe…

Cited by 66SourcePDFScholar