← Search

Xunzhi He

1 accepted papers

2026

BiasFreeBench: a Benchmark for Mitigating Bias in Large Language Model Responses

ICLR 2026poster

Existing studies on bias mitigation methods for large language models (LLMs) use diverse baselines and metrics to evaluate debiasing performance, leading to inconsistent comparisons among them. Moreover, their evaluations are mostly based on the comparison between LLMs' probabilities of biased and u…

Cited by 0SourcecodeScholar