2025
GuardAgent: Safeguard LLM Agents via Knowledge-Enabled Reasoning
ICML 2025poster
The rapid advancement of large language model (LLM) agents has raised new concerns regarding their safety and security. In this paper, we propose GuardAgent, the first guardrail agent to protect target agents by dynamically checking whether their actions satisfy given safety guard requests. Specific…