2021
Have We Learned to Explain?: How Interpretability Methods Can Learn to Encode Predictions in their Interpretations.
AISTATS 2021poster
While the need for interpretable machine learning has been established, many common approaches are slow, lack fidelity, or hard to evaluate. Amortized explanation methods reduce the cost of providing interpretations by learning a global selector model that returns feature importances for a single in…