2026
An Empirical Study of Memory Poisoning Defenses for LLM Agents
ICML 2026poster
Large Language Model (LLM) agents use memory to learn from past interactions. However, this reliance on memory introduces a critical security risk: an adversary can inject seemingly harmless records into an agent's memory to manipulate its future behavior. This vulnerability is characterized by two …