2026
Think Twice Before You Act: Enhancing Agent Behavioral Safety with Thought Correction
ICML 2026poster
LLM-based agents solve complex tasks through iterative reasoning, tool use, and environment interaction, where each intermediate thought directly shapes subsequent actions. Small deviations in these thoughts can therefore propagate into unsafe behaviors, yet existing guardrails typically operate onl…