← Search

David Haws

4 accepted papers

2024

Speak While You Think: Streaming Speech Synthesis During Text Generation

ICASSP 2024accepted

Large Language Models (LLMs) demonstrate impressive capabilities, yet interaction with these models is mostly facilitated through text. Using Text-To-Speech to synthesize LLM outputs typically results in notable latency, which is impractical for fluent voice conversations. We propose LLM2Speech, an…

Cited by 0SourceScholar
2021

Stable Checkpoint Selection and Evaluation in Sequence to Sequence Speech Synthesis

ICASSP 2021accepted

Autoregressive Attentive Sequence-to-Sequence (S2S) speech synthesis is considered state-of-the-art in terms of speech quality and naturalness, as evaluated on a finite set of testing utterances. However, it can occasionally suffer from stability issues at inference time, such as local intelligibili…

Cited by 0SourceScholar
2016

On the importance of event detection for ASR

ICASSP 2016accepted

The performance of modern large vocabulary continuous speech recognition (LVCSR) systems is heavily affected by segment boundaries, proper speaker identification of the segments, as well as removal of spurious data. We propose to use Long Short Term Memory (LSTM) recurrent neural networks to partiti…

Cited by 0SourceScholar