2026
BiasFreeBench: a Benchmark for Mitigating Bias in Large Language Model Responses
ICLR 2026poster
Existing studies on bias mitigation methods for large language models (LLMs) use diverse baselines and metrics to evaluate debiasing performance, leading to inconsistent comparisons among them. Moreover, their evaluations are mostly based on the comparison between LLMs' probabilities of biased and u…