← Search

Yanjun Fu

2 accepted papers

2025

Few-Shot Knowledge Distillation of LLMs With Counterfactual Explanations

NeurIPS 2025poster

Knowledge distillation is a promising approach to transfer capabilities from complex teacher models to smaller, resource-efficient student models that can be deployed easily, particularly in task-aware scenarios. However, existing methods of task-aware distillation typically require substantial quan…

Cited by 0SourceScholar
2025

T-SHIRT: Token-Selective Hierarchical Data Selection for Instruction Tuning

NeurIPS 2025poster

Instruction tuning is essential for Large Language Models (LLMs) to effectively follow user instructions. To improve training efficiency and reduce data redundancy, recent works use LLM-based scoring functions, e.g., Instruction-Following Difficulty (IFD), to select high–quality instruction-tuning d…

Cited by 0SourcecodeScholar