2024
LIDAO: Towards Limited Interventions for Debiasing (Large) Language Models
ICML 2024spotlight
Large language models (LLMs) have achieved impressive performance on various natural language generation tasks. Nonetheless, they suffer from generating negative and harmful contents that are biased against certain demographic groups (e.g., female), raising severe fairness concerns. As remedies, pri…