Publications-Periodical Articles

Article View/Open

Publication Export

Google ScholarTM

NCCU Library

Citation Infomation

Related Publications in TAIR

題名 Effective adversarial example detection with DeepSHAP summary
作者 郁方
Lin, Yi-Ching;Yu, Fang
貢獻者 資管系
關鍵詞 Adversarial example detection; Explainable AI; DeepSHAP; Decision logic
日期 2026-04
上傳時間 27-Aug-2026 10:32:19 (UTC+8)
摘要 Explainable AI (XAI) techniques have been widely adopted to enhance the interpretability and reliability of deep learning applications. To extend that success to adversarial example detection, we propose a new framework to extract decision logic from explanations, leverage that information to summarize common critical neurons, and utilize their status to distinguish normal and adversarial examples. Our first approach uses decision logic for detection, demonstrating that differences in critical neuron distributions can be leveraged to distinguish normal and adversarial examples. We then propose a best-layer selection strategy to enhance previous layer-wise SHAP value detection. Selecting the layer with the most common critical neurons improves performance in terms of both accuracy and computational efficiency. These two approaches achieve high detection accuracy but require runtime computation of SHAP values. To avoid such runtime overhead, we further propose a new activation status detection approach where we show that using the activation status of common critical neurons offers lightweight yet effective detection. This efficacy extends to untrained attack detection. We conduct a comprehensive study on the CIFAR-10, MNIST, SVHN, CIFAR-100, Tiny ImageNet, and ImageNet datasets to evaluate the prediction accuracy, resource consumption, and transferability of the proposed approaches against several state-of-the-art adversarial attacks. The activation status approach achieves 81.89% accuracy with the optimized parameter set, demonstrating its effectiveness and efficiency in detecting adversarial examples in high-resolution data.
關聯 Neural Computing and Applications, Vol.38, article number 278
資料類型 article
DOI https://doi.org/10.1007/s00521-026-11975-7
dc.contributor 資管系
dc.creator (作者) 郁方
dc.creator (作者) Lin, Yi-Ching;Yu, Fang
dc.date (日期) 2026-04
dc.date.accessioned 27-Aug-2026 10:32:19 (UTC+8)-
dc.date.available 27-Aug-2026 10:32:19 (UTC+8)-
dc.date.issued (上傳時間) 27-Aug-2026 10:32:19 (UTC+8)-
dc.identifier.uri (URI) https://ah.lib.nccu.edu.tw/item?item_id=184622-
dc.description.abstract (摘要) Explainable AI (XAI) techniques have been widely adopted to enhance the interpretability and reliability of deep learning applications. To extend that success to adversarial example detection, we propose a new framework to extract decision logic from explanations, leverage that information to summarize common critical neurons, and utilize their status to distinguish normal and adversarial examples. Our first approach uses decision logic for detection, demonstrating that differences in critical neuron distributions can be leveraged to distinguish normal and adversarial examples. We then propose a best-layer selection strategy to enhance previous layer-wise SHAP value detection. Selecting the layer with the most common critical neurons improves performance in terms of both accuracy and computational efficiency. These two approaches achieve high detection accuracy but require runtime computation of SHAP values. To avoid such runtime overhead, we further propose a new activation status detection approach where we show that using the activation status of common critical neurons offers lightweight yet effective detection. This efficacy extends to untrained attack detection. We conduct a comprehensive study on the CIFAR-10, MNIST, SVHN, CIFAR-100, Tiny ImageNet, and ImageNet datasets to evaluate the prediction accuracy, resource consumption, and transferability of the proposed approaches against several state-of-the-art adversarial attacks. The activation status approach achieves 81.89% accuracy with the optimized parameter set, demonstrating its effectiveness and efficiency in detecting adversarial examples in high-resolution data.
dc.format.extent 106 bytes-
dc.format.mimetype text/html-
dc.relation (關聯) Neural Computing and Applications, Vol.38, article number 278
dc.subject (關鍵詞) Adversarial example detection; Explainable AI; DeepSHAP; Decision logic
dc.title (題名) Effective adversarial example detection with DeepSHAP summary
dc.type (資料類型) article
dc.identifier.doi (DOI) 10.1007/s00521-026-11975-7
dc.doi.uri (DOI) https://doi.org/10.1007/s00521-026-11975-7