2025
Self-Debiasing Large Language Models: Zero-Shot Recognition and Reduction of Stereotypes
NAACL 2025short
Large language models (LLMs) have shown remarkable advances in language generation and understanding but are also prone to exhibiting harmful social biases. While recognition of these behaviors has generated an abundance of bias mitigation techniques, most require modifications to the training data,…