2023
PhaseAug: A Differentiable Augmentation for Speech Synthesis to Simulate One-to-Many Mapping
ICASSP 2023accepted
Previous generative adversarial network (GAN)-based neural vocoders are trained to reconstruct the exact ground truth wave-form from the paired mel-spectrogram and do not consider the one-to-many relationship of speech synthesis. This conventional training causes overfitting for both the discriminat…