Pruning Adversarially Robust Neural Networks without Adversarial Examples
Pruning Adversarially Robust Neural Networks without Adversarial Examples
复制标题
DOI:
10.1109/icdm54844.2022.00120
复制
发表时间:
2022-10
期刊:
影响因子:
--
通讯作者:
T. Jian;Zifeng Wang;Yanzhi Wang;Jennifer G. Dy;Stratis Ioannidis
中科院分区:
文献类型:
--
作者:
T. Jian;Zifeng Wang;Yanzhi Wang;Jennifer G. Dy;Stratis Ioannidis
Adversarial pruning compresses models while preserving robustness. Current methods require access to adversarial examples during pruning. This significantly hampers training efficiency. Moreover, as new adversarial attacks and training methods develop at a rapid rate, adversarial pruning methods need to be modified accordingly to keep up. In this work, we propose a novel framework to prune a previously trained robust neural network while maintaining adversarial robustness, without further generating adversarial examples. We leverage concurrent self-distillation and pruning to preserve knowledge in the original model as well as regularizing the pruned model via the Hilbert-Schmidt Information Bottleneck. We comprehensively evaluate our proposed framework and show its superior performance in terms of both adversarial robustness and efficiency when pruning architectures trained on the MNIST, CIFAR-10, and CIFAR-100 datasets against five state-of-the-art attacks..