2024
Improving LLM Attributions with Randomized Path-Integration
EMNLP 2024finding
We present Randomized Path-Integration (RPI) - a path-integration method for explaining language models via randomization of the integration path over the attention information in the model. RPI employs integration on internal attention scores and their gradients along a randomized path, which is dy…