2025
Guiding LLM Decision-Making with Fairness Reward Models
NeurIPS 2025poster
Large language models are increasingly used to support high-stakes decisions, potentially influencing who is granted bail or receives a loan. Naive chain-of-thought sampling can improve average decision accuracy, but has also been shown to amplify unfair bias. To address this challenge and enable th…