article Open access

Private Graph Extraction via Feature Explanations

  • Proceedings on Privacy Enhancing Technologies
  • De Gruyter Open
Research footprint

At a glance

Citations
8
References
49
Comments
0
Paper overview

Abstract

Privacy and interpretability are two important ingredients for achieving trustworthy machine learning. We study the interplay of these two aspects in graph machine learning through graph reconstruction attacks. The goal of the adversary here is to reconstruct the graph structure of the training data given access to model explanations. Based on the different kinds of auxiliary information available to the adversary, we propose several graph reconstruction attacks. We show that additional knowledge of post-hoc feature explanations substantially increases the success rate of these attacks. Further, we investigate in detail the differences between attack performance with respect to three different classes of explanation methods for graph neural networks: gradient-based, perturbation-based, and surrogate model-based methods. While gradient-based explanations reveal the most in terms of the graph structure, we find that these explanations do not always score high in utility. For the other two classes of explanations, privacy leakage increases with an increase in explanation utility. Finally, we propose a defense based on a randomized response mechanism for releasing the explanations, which substantially reduces the attack success rate. Our code is available at https://github.com/iyempissy/graph-stealing-attacks-with-explanation.

Record transparency

Publication details

DOI
10.56553/popets-2023-0041
OpenAlex
W4323349047
Document type
article
Language
EN
Source
Proceedings on Privacy Enhancing Technologies
Last metadata update
Community

Comments

Log in to join the discussion.

  1. No comments yet. Start the discussion.