← Search

Mohammad Rasool Izadi

4 accepted papers

2025

Simultaneous Music Separation and Generation Using Multi-Track Latent Diffusion Models

ICASSP 2025accepted

Diffusion models have recently shown strong potential in both music generation and music source separation tasks. Although in early stages, a trend is emerging towards integrating these tasks into a single framework, as both involve generating musically aligned parts and can be seen as facets of the…

Cited by 0SourceScholar
2024

"It os Okay to be Uncommon": Quantizing Sound Event Detection Networks on Hardware Accelerators with Uncommon Sub-Byte Support

ICASSP 2024accepted

If our noise-canceling headphones can understand our audio environments, they can then inform us of important sound events, tune equalization based on the types of content we listen to, and dynamically adjust noise cancellation parameters based on audio scenes to further reduce distraction. However,…

Cited by 0SourceScholar
2024

Towards Optimal Voice Disentanglement with Weak Supervision

ICASSP 2024accepted

Voice disentanglement, the process of isolating speech or singing voice into several latent subspaces, each representing certain aspects, holds significant importance in diverse audio processing applications. In this paper, we propose an efficient weakly-supervised approach to tackle this challenge.…

Cited by 0SourceScholar
2023

HiSSNet: Sound Event Detection and Speaker Identification via Hierarchical Prototypical Networks for Low-Resource Headphones

ICASSP 2023accepted

Modern noise-cancelling headphones have significantly improved users’ auditory experiences by removing unwanted background noise, but they can also block out sounds that matter to users. Machine learning (ML) models for sound event detection (SED) and speaker identification (SID) can enable headphon…

Cited by 0SourceScholar