2025
Red Queen: Exposing Latent Multi-Turn Risks in Large Language Models
ACL 2025finding
The rapid advancement of large language models (LLMs) has unlocked diverse opportunities across domains and applications but has also raised concerns about their tendency to generate harmful responses under jailbreak attacks. However, most existing jailbreak strategies are single-turn with explicit…