← Search

Yujia Yan

3 accepted papers

2024

Towards Optimal Voice Disentanglement with Weak Supervision

ICASSP 2024accepted

Voice disentanglement, the process of isolating speech or singing voice into several latent subspaces, each representing certain aspects, holds significant importance in diverse audio processing applications. In this paper, we propose an efficient weakly-supervised approach to tackle this challenge.…

Cited by 0SourceScholar
2021

Skipping the Frame-Level: Event-Based Piano Transcription With Neural Semi-CRFs

NeurIPS 2021poster

Piano transcription systems are typically optimized to estimate pitch activity at each frame of audio. They are often followed by carefully designed heuristics and post-processing algorithms to estimate note events from the frame-level predictions. Recent methods have also framed piano transcription…