← Search

Xunguang Wang

2 accepted papers

2026

GuidedBench: Measuring and Mitigating the Evaluation Discrepancies of In-the-wild LLM Jailbreak Methods

ICLR 2026poster

Despite the growing interest in jailbreaks as an effective red-teaming tool for building safe and responsible large language models (LLMs), flawed evaluation system designs have led to significant discrepancies in their effectiveness assessments. With a systematic measurement study based on 37 jailb…

Cited by 0SourcecodeScholar
2021

Prototype-Supervised Adversarial Network for Targeted Attack of Deep Hashing

CVPR 2021poster

Due to its powerful capability of representation learning and high-efficiency computation, deep hashing has made significant progress in large-scale image retrieval. However, deep hashing networks are vulnerable to adversarial examples, which is a practical secure problem but seldom studied in hashi…

Cited by 62PDFcodeScholar