2024
Understanding the Training Speedup from Sampling with Approximate Losses
ICML 2024poster
It is well known that selecting samples with large losses/gradients can significantly reduce the number of training steps. However, the selection overhead is often too high to yield any meaningful gains in terms of overall training time. In this work, we focus on the greedy approach of selecting sam…