Fine-Tune the Pretrained ATST Model for Sound Event Detection
Sound event detection (SED) often suffers from the data deficiency problem. Recent SED systems leverage the large pretrained self-supervised learning (SelfSL) models to mitigate such restriction, where the pretrained models help to produce more discriminative features for SED. However, the pretraine…