← Search

Yinzhi Zhao

1 accepted papers

2025

SemanticCamo: Jailbreaking Large Language Models through Semantic Camouflage

ACL 2025finding

The rapid development and increasingly widespread applications of Large Language Models (LLMs) have made the safety issues of LLMs more prominent and critical. Although safety training is widely used in LLMs, the mismatch between pre-training and safety training still leads to safety vulnerabilities…