← Search

Anamika Lochab

2 accepted papers

2025

VERA: Variational Inference Framework for Jailbreaking Large Language Models

NeurIPS 2025poster

The rise of API-only access to state-of-the-art LLMs highlights the need for effective black-box jailbreak methods to identify model vulnerabilities in real-world settings. Without a principled objective for gradient-based optimization, most existing approaches rely on genetic algorithms, which are…

Cited by 0SourceScholar