2023
Conceptor-Aided Debiasing of Large Language Models
EMNLP 2023long main
Pre-trained large language models (LLMs) reflect the inherent social biases of their training corpus. Many methods have been proposed to mitigate this issue, but they often fail to debias or they sacrifice model accuracy. We use *conceptors*--a soft projection method--to identify and remove the bia…