2025
Noise, Adaptation, and Strategy: Assessing LLM Fidelity in Decision-Making
EMNLP 2025
Large language models (LLMs) are increasingly used for social-science simulations, yet most evaluations target task optimality rather than the variability and adaptation characteristic of human decision-making. We propose a process-oriented evaluation framework with progressive interventions (Intrin