2024
To Each (Textual Sequence) Its Own: Improving Memorized-Data Unlearning in Large Language Models
ICML 2024poster
LLMs have been found to memorize training textual sequences and regurgitate verbatim said sequences during text generation time. This fact is known to be the cause of privacy and related (e.g., copyright) problems. Unlearning in LLMs then takes the form of devising new algorithms that will properly…