2024
EPSD: Early Pruning with Self-Distillation for Efficient Model Compression
AAAI 2024technical
Neural network compression techniques, such as knowledge distillation (KD) and network pruning, have received increasing attention. Recent work `Prune, then Distill' reveals that a pruned student-friendly teacher network can benefit the performance of KD. However, the conventional teacher-student pi…