Bandlimiting Neural Networks Against Adversarial Attacks

TOP 文献データベース Bandlimiting Neural Networks Against Adversarial Attacks

arxiv

AIセキュリティポータルbot

文献データベースの情報は、自動的に収集されています。

Source

https://arxiv.org/abs/1905.12797

PDF

https://arxiv.org/pdf/1905.12797

文献情報

作者: Yuping Lin,Kasra Ahmadi K. A.,Hui Jiang
公開日: 2019-5-30
所属機関: iFLYTEK Laboratory for Neural Computing and Machine Learning (iNCML)
所属の国: Canada
会議名: Computing Research Repository (CoRR)

AIにより推定されたラベル

深層学習ポイズニング敵対的サンプルの脆弱性

※ こちらのラベルはAIによって自動的に追加されました。そのため、正確でないことがあります。
詳細は文献データベースについてをご覧ください。

Abstract

In this paper, we study the adversarial attack and defence problem in deep learning from the perspective of Fourier analysis. We first explicitly compute the Fourier transform of deep ReLU neural networks and show that there exist decaying but non-zero high frequency components in the Fourier spectrum of neural networks. We demonstrate that the vulnerability of neural networks towards adversarial samples can be attributed to these insignificant but non-zero high frequency components. Based on this analysis, we propose to use a simple post-averaging technique to smooth out these high frequency components to improve the robustness of neural networks against adversarial attacks. Experimental results on the ImageNet dataset have shown that our proposed method is universally effective to defend many existing adversarial attacking methods proposed in the literature, including FGSM, PGD, DeepFool and C&W attacks. Our post-averaging method is simple since it does not require any re-training, and meanwhile it can successfully defend over 95% of the adversarial samples generated by these methods without introducing any significant performance degradation (less than 1%) on the original clean images.