2025
SELMA: A Speech-Enabled Language Model for Virtual Assistant Interactions
ICASSP 2025accepted
In this work, we present and evaluate SELMA, a Speech-Enabled Language Model for virtual Assistant interactions that integrates audio and text as inputs to a Large Language Model (LLM). SELMA is designed to handle three primary and two auxiliary tasks related to interactions with virtual assistants…