← Search

Ruofan Mao

1 accepted papers

2025

Chain-of-Scrutiny: Detecting Backdoor Attacks for Large Language Models

ACL 2025finding

Large Language Models (LLMs), especially those accessed via APIs, have demonstrated impressive capabilities across various domains. However, users without technical expertise often turn to (untrustworthy) third-party services, such as prompt engineering, to enhance their LLM experience, creating vul…