← Search

Kashif Munir

1 accepted papers

2025

Red Queen: Exposing Latent Multi-Turn Risks in Large Language Models

ACL 2025finding

The rapid advancement of large language models (LLMs) has unlocked diverse opportunities across domains and applications but has also raised concerns about their tendency to generate harmful responses under jailbreak attacks. However, most existing jailbreak strategies are single-turn with explicit…