← Search

Gianluigi Lopardo

2 accepted papers

2024

Attention Meets Post-hoc Interpretability: A Mathematical Perspective

ICML 2024poster

Attention-based architectures, in particular transformers, are at the heart of a technological revolution. Interestingly, in addition to helping obtain state-of-the-art results on a wide range of applications, the attention mechanism intrinsically provides meaningful insights on the internal behavio…

2023

A Sea of Words: An In-Depth Analysis of Anchors for Text Data

AISTATS 2023poster

Anchors (Ribeiro et al., 2018) is a post-hoc, rule-based interpretability method. For text data, it proposes to explain a decision by highlighting a small set of words (an anchor) such that the model to explain has similar outputs when they are present in a document. In this paper, we present the fi…