← Search

Nhi Nguyen

2 accepted papers

2026

Sufficiency is Relative: Evaluating LLM Explanations under Model-Induced Input Distributions

ICML 2026poster

Large language models (LLMs) are increasingly deployed in high-stakes domains, where free-text explanations such as chain-of-thought and post-hoc rationales are used to justify model outputs. Yet it remains unclear whether these explanations are _sufficient_, i.e., if they contain enough information…

Cited by 0SourceScholar
2024

Explanations that reveal all through the definition of encoding

NeurIPS 2024poster

Feature attributions attempt to highlight what inputs drive predictive power. Good attributions or explanations are thus those that produce inputs that retain this predictive power; accordingly, evaluations of explanations score their quality of prediction. However, evaluations produce scores better…

Cited by 1SourcePDFScholar