论文信息 - Zorro: Valid, Sparse, and Stable Explanations in Graph Neural Networks

Zorro: Valid, Sparse, and Stable Explanations in Graph Neural Networks

With the ever-increasing popularity and applications of graph neural networks, several proposals have been made to interpret and understand the decisions of a GNN model. Explanations for a GNN model differ in principle from other input settings. It is important to attribute the decision to input features and other related instances connected by the graph structure.We find that the previous explanation generation approaches that maximize the mutual information between the label distribution produced by the GNN model and the explanation to be restrictive. Specifically, existing approaches do not enforce explanations to be predictive, sparse, or robust to input perturbations. In this paper, we lay down some of the fundamental principles that an explanation method for GNNs should follow and introduce a metric fidelity as a measure of the explanation’s effectiveness. We propose a novel approach Zorro based on the principles from rate-distortion theory that uses a simple combinatorial procedure to optimize for fidelity. Extensive experiments on real and synthetic datasets reveal that Zorro produces sparser, stable, andmore faithful explanations than existing GNN explanation approaches. ACM Reference Format: Thorben Funke, Megha Khosla, and Avishek Anand. 2018. Zorro: Valid, Sparse, and Stable Explanations in Graph Neural Networks. InWoodstock ’18: ACM Symposium on Neural Gaze Detection, June 03–05, 2018, Woodstock, NY . ACM,NewYork, NY, USA, 12 pages. https://doi.org/10.1145/1122445.1122456

Megha Khosla | Avishek Anand | Thorben Funke

[1] Avanti Shrikumar,et al. Learning Important Features Through Propagating Activation Differences , 2017, ICML.

[2] Jonas Mueller,et al. What made you do this? Understanding black-box decisions with sufficient input subsets , 2018, AISTATS.

[3] Le Song,et al. Learning to Explain: An Information-Theoretic Perspective on Model Interpretation , 2018, ICML.

[4] Dumitru Erhan,et al. A Benchmark for Interpretability Methods in Deep Neural Networks , 2018, NeurIPS.

[5] Max Welling,et al. Semi-Supervised Classification with Graph Convolutional Networks , 2016, ICLR.

[6] Lior Wolf,et al. A Formal Approach to Explainability , 2019, AIES.

[7] Ruslan Salakhutdinov,et al. Revisiting Semi-Supervised Learning with Graph Embeddings , 2016, ICML.

[8] My T. Thai,et al. PGM-Explainer: Probabilistic Graphical Model Explanations for Graph Neural Networks , 2020, NeurIPS.

[9] Megha Khosla,et al. Finding Interpretable Concept Spaces in Node Embeddings using Knowledge Bases , 2019, PKDD/ECML Workshops.

[10] Steven Skiena,et al. DeepWalk: online learning of social representations , 2014, KDD.

[11] Dumitru Erhan,et al. The (Un)reliability of saliency methods , 2017, Explainable AI.