Sanitizing hidden activations for improving adversarial robustness of convolutional neural networks

Tianshi Mu,Jian Wang,Kequan Lin,Huabing Zhang

doi:10.3233/jifs-210371

Abstract

Deep learning is gaining significant traction in a wide range of areas. Whereas, recent studies have demonstrated that deep learning exhibits the fatal weakness on adversarial examples. Due to the black-box nature and un-transparency problem of deep learning, it is difficult to explain the reason for the existence of adversarial examples and also hard to defend against them. This study focuses on improving the adversarial robustness of convolutional neural networks. We first explore how adversarial examples behave inside the network through visualization. We find that adversarial examples produce perturbations in hidden activations, which forms an amplification effect to fool the network. Motivated by this observation, we propose an approach, termed as sanitizing hidden activations, to help the network correctly recognize adversarial examples by eliminating or reducing the perturbations in hidden activations. To demonstrate the effectiveness of our approach, we conduct experiments on three widely used datasets: MNIST, CIFAR-10 and ImageNet, and also compare with state-of-the-art defense techniques. The experimental results show that our sanitizing approach is more generalized to defend against different kinds of attacks and can effectively improve the adversarial robustness of convolutional neural networks.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Sanitizing hidden activations for improving adversarial robustness of convolutional neural networks

Abstract

Talk to us

Similar Papers

More From: Journal of Intelligent & Fuzzy Systems

Lead the way for us

Similar Papers

Assessment of the Robustness of Convolutional Neural Networks in Labeling Noise by Using Chest X-Ray Images From Multiple Centers.
Ryoungwoo Jang ... Namkug Kim
JMIR Medical Informatics | VOL. 8
Ryoungwoo Jang, et. al.Ryoungwoo Jang ... Namkug Kim
04 Aug 2020
JMIR Medical Informatics | VOL. 8

Adversarial Examples in Deep Neural Networks: An Overview
Emilio Rafael Balda ... Rudolf Mathar
-
Emilio Rafael Balda, et. al.Emilio Rafael Balda ... Rudolf Mathar
24 Oct 2019
24 Oct 2019

Dealing with Robustness of Convolutional Neural Networks for Image Classification
Paolo Arcaini ... Andrea Bombarda
-
Paolo Arcaini, et. al.Paolo Arcaini ... Andrea Bombarda
01 Aug 2020
01 Aug 2020

Sparse adversarial image generation using dictionary learning
Maham Jahangir ... Faisal Shafait
Journal of Electronic Imaging | VOL. 32
Maham Jahangir, et. al.Maham Jahangir ... Faisal Shafait
18 May 2023
Journal of Electronic Imaging | VOL. 32

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Sanitizing hidden activations for improving adversarial robustness of convolutional neural networks

Abstract

Talk to us

Similar Papers

More From: Journal of Intelligent &amp; Fuzzy Systems

More From: Journal of Intelligent & Fuzzy Systems