← Search

Jacy Reese Anthis

4 accepted papers

2025

Bias in Language Models: Beyond Trick Tests and Towards RUTEd Evaluation

ACL 2025long

Standard bias benchmarks used for large language models (LLMs) measure the association between social attributes in model inputs and single-word model outputs. We test whether these benchmarks are robust to lengthening the model outputs via a more realistic user prompt, in the commonly studied domai…

Cited by 0SourcePDFScholar
2025

Position: LLM Social Simulations Are a Promising Research Method

ICML 2025poster

Accurate and verifiable large language model (LLM) simulations of human research subjects promise an accessible data source for understanding human behavior and training new AI systems. However, results to date have been limited, and few social scientists have adopted this method. In this position p…

Cited by 4SourcePDFScholar
2023

Causal Context Connects Counterfactual Fairness to Robust Prediction and Group Fairness

NeurIPS 2023poster

Counterfactual fairness requires that a person would have been classified in the same way by an AI or other algorithmic system if they had a different protected class, such as a different race or gender. This is an intuitive standard, as reflected in the U.S. legal system, but its use is limited bec…