2025
MoWE-Audio: Multitask AudioLLMs with Mixture of Weak Encoders
ICASSP 2025accepted
The rapid advancements in large language models (LLMs) have significantly enhanced natural language processing capabilities, facilitating the development of AudioLLMs that process and understand speech and audio inputs alongside text. Existing AudioLLMs typically combine a pre-trained audio encoder…