Attacks Meet Interpretability: Attribute-steered Detection of Adversarial Samples

Guanhong Tao, ShiQing Ma, Yingqi Liu, Xiangyu Zhang. Attacks Meet Interpretability: Attribute-steered Detection of Adversarial Samples. In Samy Bengio, Hanna M. Wallach, Hugo Larochelle, Kristen Grauman, Nicolò Cesa-Bianchi, Roman Garnett, editors, Advances in Neural Information Processing Systems 31: Annual Conference on Neural Information Processing Systems 2018, NeurIPS 2018, 3-8 December 2018, Montréal, Canada. pages 7728-7739, 2018. [doi]

Abstract

Abstract is missing.