Adversarial Neural Network Inversion via Auxiliary Knowledge Alignment

TOP 文献データベース Adversarial Neural Network Inversion via Auxiliary Knowledge Alignment

Computing Research Repository (CoRR)

AIセキュリティポータルbot

文献データベースの情報は、自動的に収集されています。

Source

https://arxiv.org/abs/1902.08552

PDF

https://arxiv.org/pdf/1902.08552

文献情報

作者: Ziqi Yang,Ee-Chien Chang,Zhenkai Liang
公開日: 2019-2-23
所属機関: School of Computing, National University of Singapore
所属の国: Singapore
会議名: Computing Research Repository (CoRR)

AIにより推定されたラベル

モデルインバージョン最適化手法敵対的攻撃手法

※ こちらのラベルはAIによって自動的に追加されました。そのため、正確でないことがあります。
詳細は文献データベースについてをご覧ください。

Abstract

The rise of deep learning technique has raised new privacy concerns about the training data and test data. In this work, we investigate the model inversion problem in the adversarial settings, where the adversary aims at inferring information about the target model's training data and test data from the model's prediction values. We develop a solution to train a second neural network that acts as the inverse of the target model to perform the inversion. The inversion model can be trained with black-box accesses to the target model. We propose two main techniques towards training the inversion model in the adversarial settings. First, we leverage the adversary's background knowledge to compose an auxiliary set to train the inversion model, which does not require access to the original training data. Second, we design a truncation-based technique to align the inversion model to enable effective inversion of the target model from partial predictions that the adversary obtains on victim user's data. We systematically evaluate our inversion approach in various machine learning tasks and model architectures on multiple image datasets. Our experimental results show that even with no full knowledge about the target model's training data, and with only partial prediction values, our inversion approach is still able to perform accurate inversion of the target model, and outperform previous approaches.