← Search

Manuel Costa

1 accepted papers

2026

Optimizing Agent Planning for Security and Autonomy

ICLR 2026poster

Indirect prompt injection attacks threaten AI agents that execute consequential actions, motivating deterministic system-level defenses. Such defenses can provably block unsafe actions by enforcing confidentiality and integrity policies, but currently appear costly: they reduce task completion rates…

Cited by 0SourcecodeScholar