← Search

Damien Garreau

14 accepted papers

2024

Attention Meets Post-hoc Interpretability: A Mathematical Perspective

ICML 2024poster

Attention-based architectures, in particular transformers, are at the heart of a technological revolution. Interestingly, in addition to helping obtain state-of-the-art results on a wide range of applications, the attention mechanism intrinsically provides meaningful insights on the internal behavio…

2023

A Sea of Words: An In-Depth Analysis of Anchors for Text Data

AISTATS 2023poster

Anchors (Ribeiro et al., 2018) is a post-hoc, rule-based interpretability method. For text data, it proposes to explain a decision by highlighting a small set of words (an anchor) such that the model to explain has similar outputs when they are present in a document. In this paper, we present the fi…

2023

Explainability as statistical inference

ICML 2023poster

A wide variety of model explanation approaches have been proposed in recent years, all guided by very different rationales and heuristics. In this paper, we take a new route and cast interpretability as a statistical inference problem. We propose a general deep probabilistic model designed to produc…

Cited by 5SourcePDFScholar