← Search

Edoardo Pona

1 accepted papers

2025

Abstract Counterfactuals for Language Model Agents

NeurIPS 2025poster

Counterfactual inference is a powerful tool for analysing and evaluating autonomous agents, but its application to language model (LM) agents remains challenging. Existing work on counterfactuals in LMs has primarily focused on token-level counterfactuals, which are often inadequate for LM agents du…

Cited by 0SourceScholar