← Search

Mehrnaz Mofakhami

4 accepted papers

2026

A Coin Flip for Safety: LLM Judges Fail to Reliably Measure Adversarial Robustness

ICML 2026poster

Automated \enquote{LLM-as-a-Judge} frameworks have become the de facto standard for scalable evaluation across natural language processing. For instance, in safety evaluation, these judges are relied upon to evaluate harmfulness in order to benchmark the robustness of safety against adversarial atta…

Cited by 0SourceScholar
2025

Performative Prediction on Games and Mechanism Design

AISTATS 2025poster

Agents often have individual goals which depend on a group's actions. If agents trust a forecast of collective action and adapt strategically, such prediction can influence outcomes non-trivially, resulting in a form of performative prediction. This effect is ubiquitous in scenarios ranging from pan…

Cited by 0SourcecodeScholar
2025

Tight Lower Bounds and Improved Convergence in Performative Prediction

NeurIPS 2025poster

Performative prediction is a framework accounting for the shift in the data distribution induced by the prediction of a model deployed in the real world. Ensuring convergence to a stable solution—one at which the post‑deployment data distribution no longer changes—is crucial in settings where model…

Cited by 0SourcecodeScholar