← Search

Hendrik Schuff

3 accepted papers

2024

Explaining Pre-Trained Language Models with Attribution Scores: An Analysis in Low-Resource Settings

COLING 2024main

Attribution scores indicate the importance of different input parts and can, thus, explain model behaviour. Currently, prompt-based models are gaining popularity, i.a., due to their easier adaptability in low-resource settings. However, the quality of attribution scores extracted from prompt-based m…

Cited by 2SourcePDFScholar
2023

Neighboring Words Affect Human Interpretation of Saliency Explanations

ACL 2023findings

Word-level saliency explanations (“heat maps over words”) are often used to communicate feature-attribution in text-based models. Recent studies found that superficial factors such as word length can distort human interpretation of the communicated saliency scores. We conduct a user study to investi…