← Search

Max Mathys

2 accepted papers

2026

Breaking Agent Backbones: Evaluating the Security of Backbone LLMs in AI Agents

ICLR 2026poster

AI agents powered by large language models (LLMs) are being deployed at scale, yet we lack a systematic understanding of how the choice of backbone LLM affects agent security. The non-deterministic sequential nature of AI agents complicates security modeling, while the integration of traditional sof…

Cited by 0SourceScholar
2025

Gandalf the Red: Adaptive Security for LLMs

ICML 2025poster

Current evaluations of defenses against prompt attacks in large language model (LLM) applications often overlook two critical factors: the dynamic nature of adversarial behavior and the usability penalties imposed on legitimate users by restrictive defenses. We propose D-SEC (Dynamic Security Utilit…