Sanitizing hidden activations for improving adversarial robustness of convolutional neural networks

被引：0

作者：

Mu, Tianshi ^{[1
]}

Lin, Kequan ^{[1
]}

Zhang, Huabing ^{[1
]}

Wang, Jian ^{[1
]}

机构：

[1] China Southern Power Grid, Digital Grid Res Inst, Guangzhou 510700, Peoples R China

来源：

JOURNAL OF INTELLIGENT & FUZZY SYSTEMS | 2021年 / 41卷 / 02期

关键词：

Adversarial examples; sanitizing hidden activations; adversarial robustness; convolutional neural networks;

D O I：

10.3233/JIFS-210371

中图分类号：

TP18 [人工智能理论];

学科分类号：

081104 ; 0812 ; 0835 ; 1405 ;

摘要：

Deep learning is gaining significant traction in a wide range of areas. Whereas, recent studies have demonstrated that deep learning exhibits the fatal weakness on adversarial examples. Due to the black-box nature and un-transparency problem of deep learning, it is difficult to explain the reason for the existence of adversarial examples and also hard to defend against them. This study focuses on improving the adversarial robustness of convolutional neural networks. We first explore how adversarial examples behave inside the network through visualization. We find that adversarial examples produce perturbations in hidden activations, which forms an amplification effect to fool the network. Motivated by this observation, we propose an approach, termed as sanitizing hidden activations, to help the network correctly recognize adversarial examples by eliminating or reducing the perturbations in hidden activations. To demonstrate the effectiveness of our approach, we conduct experiments on three widely used datasets: MNIST, CIFAR-10 and ImageNet, and also compare with state-of-the-art defense techniques. The experimental results show that our sanitizing approach is more generalized to defend against different kinds of attacks and can effectively improve the adversarial robustness of convolutional neural networks.

引用

页码：3993 / 4003

页数：11

共 50 条

[31] A SIMPLE STOCHASTIC NEURAL NETWORK FOR IMPROVING ADVERSARIAL ROBUSTNESS
Yang, Hao
Wang, Min
Yu, Zhengfei
Zhou, Yun
[J]. 2023 IEEE INTERNATIONAL CONFERENCE ON MULTIMEDIA AND EXPO, ICME, 2023, : 2297 - 2302
[32] Convolutional and generative adversarial neural networks in manufacturing
Kusiak, Andrew
[J]. INTERNATIONAL JOURNAL OF PRODUCTION RESEARCH, 2020, 58 (05) : 1594 - 1604
[33] A training strategy for improving the robustness of memristor-based binarized convolutional neural networks
Huang, Lixing
Yu, Hongqi
Chen, Changlin
Peng, Jie
Diao, Jietao
Nie, Hongshan
Li, Zhiwei
Liu, Haijun
[J]. SEMICONDUCTOR SCIENCE AND TECHNOLOGY, 2022, 37 (01)
[34] Characterizing Adversarial Samples of Convolutional Neural Networks
Jiang, Cheng
Zhao, Qiyang
Liu, Yuzhong
[J]. 2018 11TH INTERNATIONAL CONGRESS ON IMAGE AND SIGNAL PROCESSING, BIOMEDICAL ENGINEERING AND INFORMATICS (CISP-BMEI 2018), 2018,
[35] Improving robustness of convolutional neural networks using element-wise activation scaling
Zhang, Zhi-Yuan
Ren, Hao
He, Zhenli
Zhou, Wei
Liu, Di
[J]. FUTURE GENERATION COMPUTER SYSTEMS-THE INTERNATIONAL JOURNAL OF ESCIENCE, 2023, 149 : 136 - 148
[36] Efficient Computation of Robustness of Convolutional Neural Networks
Arcaini, Paolo
Bombarda, Andrea
Bonfanti, Silvia
Gargantini, Angelo
[J]. THIRD IEEE INTERNATIONAL CONFERENCE ON ARTIFICIAL INTELLIGENCE TESTING (AITEST 2021), 2021, : 21 - 28
[37] Improving Error Related Potential Classification by using Generative Adversarial Networks and Deep Convolutional Neural Networks
Gao, Chenguang
Li, Zhao
Ora, Hiroki
Miyake, Yoshihiro
[J]. 2020 IEEE INTERNATIONAL CONFERENCE ON BIOINFORMATICS AND BIOMEDICINE, 2020, : 2468 - 2476
[38] Integer Convolutional Neural Networks with Boolean Activations: The BoolHash Algorithm
Gatchev, Grigor
Mollov, Valentin S.
[J]. 24TH IEEE EUROPEAN CONFERENCE ON CIRCUIT THEORY AND DESIGN (ECCTD 2020), 2020,
[39] Training Deep Photonic Convolutional Neural Networks With Sinusoidal Activations
Passalis, Nikolaos
Mourgias-Alexandris, George
Tsakyridis, Apostolos
Pleros, Nikos
Tefas, Anastasios
[J]. IEEE TRANSACTIONS ON EMERGING TOPICS IN COMPUTATIONAL INTELLIGENCE, 2021, 5 (03): : 384 - 393
[40] Improving the Robustness of Neural Networks Using K-Support Norm Based Adversarial Training
Akhtar, Sheikh Waqas
Rehman, Saad
Akhtar, Mahmood
Khan, Muazzam A.
Riaz, Farhan
Chaudry, Qaiser
Young, Rupert
[J]. IEEE ACCESS, 2016, 4 : 9501 - 9511

← 1 2 3 4 5 →