NeurIPS 2022accept11 citations

Explaining Preferences with Shapley Values

Robert Hu, Siu Lun Chau, Jaime Ferrando Huertas, Dino Sejdinovic

Abstract

While preference modelling is becoming one of the pillars of machine learning, the problem of preference explanation remains challenging and underexplored. In this paper, we propose \textsc{Pref-SHAP}, a Shapley value-based model explanation framework for pairwise comparison data. We derive the appropriate value functions for preference models and further extend the framework to model and explain \emph{context specific} information, such as the surface type in a tennis game. To demonstrate the utility of \textsc{Pref-SHAP}, we apply our method to a variety of synthetic and real-world datasets and show that richer and more insightful explanations can be obtained over the baseline.

InterpretabilityPreference LearningKernelShapley ValuesRKHS
BibTeX
@inproceedings{
hu2022explaining,
title={Explaining Preferences with Shapley Values},
author={Robert Hu and Siu Lun Chau and Jaime Ferrando Huertas and Dino Sejdinovic},
booktitle={Advances in Neural Information Processing Systems},
editor={Alice H. Oh and Alekh Agarwal and Danielle Belgrave and Kyunghyun Cho},
year={2022},
url={https://openreview.net/forum?id=-me36V0os8P}
}