NeurIPS 2025oral0 citations

GnnXemplar: Exemplars to Explanations - Natural Language Rules for Global GNN Interpretability

Burouj Armgaan, Eshan Jain, Harsh Pandey, Mahesh Chandran, Sayan Ranu

Abstract

Graph Neural Networks (GNNs) are widely used for node classification, yet their opaque decision-making limits trust and adoption. While local explanations offer insights into individual predictions, global explanation methods—those that characterize an entire class—remain underdeveloped. Existing global explainers rely on motif discovery in small graphs, an approach that breaks down in large, real-world settings where subgraph repetition is rare, node attributes are high-dimensional, and predictions arise from complex structure-attribute interactions. We propose GnnXemplar, a novel global explainer inspired from Exemplar Theory from cognitive science. GnnXemplar identifies representative nodes in the GNN embedding space—exemplars—and explains predictions using natural language rules derived from their neighborhoods. Exemplar selection is framed as a coverage maximization problem over reverse $k$-nearest neighbors, for which we provide an efficient greedy approximation. To derive interpretable rules, we employ a self-refining prompt strategy using large language models (LLMs). Experiments across diverse benchmarks show that GnnXemplar significantly outperforms existing methods in fidelity, scalability, and human interpretability, as validated by a user study with 60 participants.

graph neural networkgraph machine learningexplainabilityxaiglobal explanationtext-based explanationexemplarexemplar theory
BibTeX
@inproceedings{
armgaan2025gnnxemplar,
title={GnnXemplar: Exemplars to Explanations - Natural Language Rules for Global {GNN} Interpretability},
author={Burouj Armgaan and Eshan Jain and Harsh Pandey and Mahesh Chandran and Sayan Ranu},
booktitle={The Thirty-ninth Annual Conference on Neural Information Processing Systems},
year={2025},
url={https://openreview.net/forum?id=eafIjoZAHm}
}
GnnXemplar: Exemplars to Explanations - Natural Language Rules for Global GNN Interpretability · NeurIPS 2025