Reaching Data Confidentiality and Model Accountability on the CalTrain

TOP 文献データベース Reaching Data Confidentiality and Model Accountability on the CalTrain

arxiv

AIセキュリティポータルbot

文献データベースの情報は、自動的に収集されています。

Source

https://arxiv.org/abs/1812.03230

PDF

https://arxiv.org/pdf/1812.03230

文献情報

作者: Zhongshu Gu,Hani Jamjoom,Dong Su,Heqing Huang,Jialong Zhang,Tengfei Ma,Dimitrios Pendarakis,Ian Molloy
公開日: 2018-12-8
所属機関: IBM Research
所属の国: United States of America
会議名: Dependable Systems and Networks (DSN)

AIにより推定されたラベル

トリガーの検知連合学習パフォーマンス評価

※ こちらのラベルはAIによって自動的に追加されました。そのため、正確でないことがあります。
詳細は文献データベースについてをご覧ください。

Abstract

Distributed collaborative learning (DCL) paradigms enable building joint machine learning models from distrusting multi-party participants. Data confidentiality is guaranteed by retaining private training data on each participant's local infrastructure. However, this approach to achieving data confidentiality makes today's DCL designs fundamentally vulnerable to data poisoning and backdoor attacks. It also limits DCL's model accountability, which is key to backtracking the responsible "bad" training data instances/contributors. In this paper, we introduce CALTRAIN, a Trusted Execution Environment (TEE) based centralized multi-party collaborative learning system that simultaneously achieves data confidentiality and model accountability. CALTRAIN enforces isolated computation on centrally aggregated training data to guarantee data confidentiality. To support building accountable learning models, we securely maintain the links between training instances and their corresponding contributors. Our evaluation shows that the models generated from CALTRAIN can achieve the same prediction accuracy when compared to the models trained in non-protected environments. We also demonstrate that when malicious training participants tend to implant backdoors during model training, CALTRAIN can accurately and precisely discover the poisoned and mislabeled training data that lead to the runtime mispredictions.