← Search

Shiyang Wang

1 accepted papers

2024

LIDAO: Towards Limited Interventions for Debiasing (Large) Language Models

ICML 2024spotlight

Large language models (LLMs) have achieved impressive performance on various natural language generation tasks. Nonetheless, they suffer from generating negative and harmful contents that are biased against certain demographic groups (e.g., female), raising severe fairness concerns. As remedies, pri…

Cited by 0SourcePDFScholar