← Search

Mateo Rojas-Carulla

3 accepted papers

2026

Breaking Agent Backbones: Evaluating the Security of Backbone LLMs in AI Agents

ICLR 2026poster

AI agents powered by large language models (LLMs) are being deployed at scale, yet we lack a systematic understanding of how the choice of backbone LLM affects agent security. The non-deterministic sequential nature of AI agents complicates security modeling, while the integration of traditional sof…

Cited by 0SourceScholar
2025

Gandalf the Red: Adaptive Security for LLMs

ICML 2025poster

Current evaluations of defenses against prompt attacks in large language model (LLM) applications often overlook two critical factors: the dynamic nature of adversarial behavior and the usability penalties imposed on legitimate users by restrictive defenses. We propose D-SEC (Dynamic Security Utilit…

2018

Learning Independent Causal Mechanisms

ICML 2018oral

Statistical learning relies upon data sampled from a distribution, and we usually do not care what actually generated it in the first place. From the point of view of causal modeling, the structure of each distribution is induced by physical mechanisms that give rise to dependences between observabl…

Cited by 206SourcePDFScholar