← Search

Sung-Feng Huang

3 accepted papers

2026

How Does Instrumental Music Help SingFake Detection?

ICASSP 2026poster

Although many models exist to detect singing voice deepfakes (SingFake), how these models operate, particularly with instrumental accompaniment, is unclear. We investigate how instrumental music affects SingFake detection from two perspectives. To investigate the behavioral effect, we test different…

Cited by 0SourcePDFScholar
2025

Generative Speech Foundation Model Pretraining for High-Quality Speech Extraction and Restoration

ICASSP 2025accepted

This paper proposes a generative pretraining foundation model for high-quality speech restoration tasks. By directly operating on complex-valued short-time Fourier transform coefficients, our model does not rely on any vocoders for time-domain signal reconstruction. As a result, our model simplifies…

Cited by 0SourceScholar
2023

Personalized Lightweight Text-to-Speech: Voice Cloning with Adaptive Structured Pruning

ICASSP 2023accepted

Personalized TTS is an exciting and highly desired application that allows users to train their TTS voice using only a few recordings. However, TTS training typically requires many hours of recording and a large model, making it unsuitable for deployment on mobile devices. To overcome this limitatio…

Cited by 0SourceScholar