2024
Ethos: Rectifying Language Models in Orthogonal Parameter Space
NAACL 2024findings
Language models (LMs) have greatly propelled the research on natural language processing. However, LMs also raise concerns regarding the generation of biased or toxic content and the potential disclosure of private information from the training dataset. In this work, we present a new efficient appro…