← Search

Weiji Zhuang

3 accepted papers

2026

Flow2GAN: Hybrid Flow Matching and GAN with Multi-Resolution Network for One-/Two-step High-Fidelity Audio Generation

ICLR 2026poster

Existing dominant methods for audio generation include Generative Adversarial Networks (GANs) and diffusion-based methods like Flow Matching. GANs suffer from slow convergence and potential mode collapse during training, while diffusion methods require multi-step inference that introduces considerab…

Cited by 0SourcecodeScholar
2022

Learning Decoupling Features Through Orthogonality Regularization

ICASSP 2022accepted

Keyword spotting (KWS) and speaker verification (SV) are two important tasks in speech applications. Research shows that the state-of-art KWS and SV models are trained independently using different datasets since they expect to learn distinctive acoustic features. However, humans can distinguish lan…

Cited by 0SourceScholar
2021

AutoKWS: Keyword Spotting with Differentiable Architecture Search

ICASSP 2021accepted

Smart audio devices are gated by an always-on lightweight keyword spotting program to reduce power consumption. It is however challenging to design models that have both high accuracy and low latency for accurate and fast responsiveness. Many efforts have been made to develop end-to-end neural netwo…

Cited by 0SourceScholar