2025
KANGAN-AVSS: Kolmogorov-Arnold Network Based Generative Adversarial Networks for Audio-Visual Speech Synthesis
ICASSP 2025accepted
Audio-visual speech synthesis (AVSS) is an emerging research topic in the paradigm of generative AI, aiming to generate realistic and synchronized audio-visual outputs for a target speaker based on input audio from any source speaker, combining Voice Conversion (VC) and Audio-Visual Synthesis (AVS).…