← Search

Jacob Goldman-Wetzler

1 accepted papers

2025

Distillation Robustifies Unlearning

NeurIPS 2025spotlight

Current LLM unlearning methods are not robust. A few steps of finetuning can revert their effects. We begin by showing that this is true even for an idealized form of unlearning: training to imitate a model that was never trained on unwanted information. This shows that training a model can drastica…

Cited by 0SourceScholar