← Search

Karl El Hajal

3 accepted papers

2025

kNN Retrieval for Simple and Effective Zero-Shot Multi-speaker Text-to-Speech

NAACL 2025short

While recent zero-shot multi-speaker text-to-speech (TTS) models achieve impressive results, they typically rely on extensive transcribed speech datasets from numerous speakers and intricate training pipelines. Meanwhile, self-supervised learning (SSL) speech features have emerged as effective inter…

2023

Efficient Speech Quality Assessment Using Self-Supervised Framewise Embeddings

ICASSP 2023accepted

Automatic speech quality assessment is essential for audio researchers, developers, speech and language pathologists, and system quality engineers. The current state-of-the-art systems are based on framewise speech features (hand-engineered or learnable) combined with time dependency modeling. This…

Cited by 0SourceScholar
2023

Personalized Task Load Prediction in Speech Communication

ICASSP 2023accepted

Estimating the quality of remote speech communication is a complex task influenced by the speaker, transmission channel, and listener. For example, the degradation of transmission quality can increase listeners’ cognitive load, which can influence the overall perceived quality of the conversation. T…

Cited by 0SourceScholar