2025
ICLShield: Exploring and Mitigating In-Context Learning Backdoor Attacks
ICML 2025poster
In-context learning (ICL) has demonstrated remarkable success in large language models (LLMs) due to its adaptability and parameter-free nature. However, it also introduces a critical vulnerability to backdoor attacks, where adversaries can manipulate LLM behaviors by simply poisoning a few ICL demo…