2024
Investigating the Impact of Model Instability on Explanations and Uncertainty
ACL 2024findings
Explainable AI methods facilitate the understanding of model behaviour, yet, small, imperceptible perturbations to inputs can vastly distort explanations. As these explanations are typically evaluated holistically, before model deployment, it is difficult to assess when a particular explanation is t…