← Search

Sambal Shikhar

1 accepted papers

2025

LLMVoX: Autoregressive Streaming Text-to-Speech Model for Any LLM

ACL 2025finding

Recent advancements in speech-to-speech dialogue systems leverage LLMs for multimodal interactions, yet they remain hindered by fine-tuning requirements, high computational overhead, and text-speech misalignment. Existing speech-enabled LLMs often degrade conversational quality by modifying the LLM,…