2026
Preventing Robotic Jailbreaking Via Multimodal Domain Adaptation
Francesco Marchiori, Rohan Sinha, Christopher George Agia, Alexander Robey, George J. Pappas, Mauro Conti +1
ICRA 2026poster
Large Language Models (LLMs) and Vision-Language Models (VLMs) are increasingly deployed in robotic environments but remain vulnerable to jailbreaking attacks that bypass safety mechanisms and drive unsafe or physically harmful behaviors in the real world. Data-driven defenses such as jailbreak clas…