2026
Speech-Audio Compositional Attacks on Multimodal LLMs and Their Defense with SALMONN-Guard
ICML 2026poster
Recent progress in large language models (LLMs) has enabled understanding of both speech and non-speech audio, but has also exposed new safety risks arising from complex audio inputs that are inadequately handled by current safeguards. We introduce SACRED-Bench (Speech–Audio Composition for RED-team…