2025
ALPACA AGAINST VICUNA: Using LLMs to Uncover Memorization of LLMs
NAACL 2025long
In this paper, we investigate the overlooked impact of instruction-tuning on memorization in large language models (LLMs), which has largely been studied in base, pre-trained models. We propose a black-box prompt optimization method where an attacker LLM agent uncovers higher levels of memorization…