2025
SemanticCamo: Jailbreaking Large Language Models through Semantic Camouflage
ACL 2025finding
The rapid development and increasingly widespread applications of Large Language Models (LLMs) have made the safety issues of LLMs more prominent and critical. Although safety training is widely used in LLMs, the mismatch between pre-training and safety training still leads to safety vulnerabilities…