2026
Reinforcement Unlearning via Group Relative Policy Optimization
ICLR 2026poster
During pretraining, LLMs inadvertently memorize sensitive or copyrighted data, posing significant compliance challenges under legal frameworks like the GDPR and the EU AI Act. Fulfilling these mandates demands techniques that can remove information from a deployed model without retraining from scrat…