← Search

Tongyu Wen

1 accepted papers

2025

Defending against Indirect Prompt Injection by Instruction Detection

EMNLP 2025

The integration of Large Language Models (LLMs) with external sources is becoming increasingly common, with Retrieval-Augmented Generation (RAG) being a prominent example. However, this integration introduces vulnerabilities of Indirect Prompt Injection (IPI) attacks, where hidden instructions embed