2025
McBE: A Multi-task Chinese Bias Evaluation Benchmark for Large Language Models
ACL 2025finding
As large language models (LLMs) are increasingly applied to various NLP tasks, their inherent biases are gradually disclosed. Therefore, measuring biases in LLMs is crucial to mitigate its ethical risks. However, most existing bias evaluation datasets are focus on English andNorth American culture,…