← Search

Seungwon Jeong

1 accepted papers

2026

SlotGCG: Exploiting the Positional Vulnerability in LLMs for Jailbreak Attacks

ICLR 2026poster

As large language models (LLMs) are widely deployed, identifying their vulnerability through jailbreak attacks becomes increasingly critical. Optimization-based attacks like Greedy Coordinate Gradient (GCG) have focused on inserting adversarial tokens to the end of prompts. However, GCG restricts ad…

Cited by 0SourcecodeScholar