Can We "Cherry-Pick"? Investigating Multiple Renditions from a Generative Speech Synthesis Model
Generative Speech Models (GSMs) have seen a surge in popularity due to their ability to generate diverse and high-quality speech. Evaluating models that generate many different renditions for a given input sentence presents a new challenge. Listening tests are still the gold standard for evaluating…