Boosting Speech Enhancement with Clean Self-Supervised Features Via Conditional Variational Autoencoders
Recently, Self-Supervised Features (SSF) trained on extensive speech datasets have shown significant performance gains across various speech processing tasks. Nevertheless, their effectiveness in Speech Enhancement (SE) systems is often suboptimal due to insufficient optimization for noisy environme…