Skip to main navigation Skip to search Skip to main content

Explaining neural networks using attentive knowledge distillation

  • Hyeonseok Lee
  • , Sungchan Kim*
  • *Corresponding author for this work
  • Jeonbuk National University

Research output: Contribution to journalJournal articlepeer-review

Abstract

Explaining the prediction of deep neural networks makes the networks more understandable and trusted, leading to their use in various mission critical tasks. Recent progress in the learning capability of networks has primarily been due to the enormous number of model parameters, so that it is usually hard to interpret their operations, as opposed to classical white-box models. For this purpose, generating saliency maps is a popular approach to identify the important input features used for the model prediction. Existing explanation methods typically only use the output of the last convolution layer of the model to generate a saliency map, lacking the information included in intermediate layers. Thus, the corresponding explanations are coarse and result in limited accuracy. Although the accuracy can be improved by iteratively developing a saliency map, this is too timeconsuming and is thus impractical. To address these problems, we proposed a novel approach to explain the model prediction by developing an attentive surrogate network using the knowledge distillation. The surrogate network aims to generate a fine-grained saliency map corresponding to the model prediction using meaningful regional information presented over all network layers. Experiments demonstrated that the saliency maps are the result of spatially attentive features learned from the distillation. Thus, they are useful for fine-grained classification tasks. Moreover, the proposed method runs at the rate of 24.3 frames per second, which is much faster than the existing methods by orders of magnitude.

Original languageEnglish
Article number1280
Pages (from-to)1-17
Number of pages17
JournalSensors
Volume21
Issue number4
DOIs
StatePublished - 2021.02.2

Keywords

  • Attention
  • Deep neural networks
  • Fine-grained classification
  • Knowledge distillation
  • Visual explanation

Quacquarelli Symonds(QS) Subject Topics

  • Computer Science & Information Systems
  • Engineering - Electrical & Electronic
  • Engineering - Petroleum
  • Chemistry
  • Physics & Astronomy
  • Biological Sciences

Fingerprint

Dive into the research topics of 'Explaining neural networks using attentive knowledge distillation'. Together they form a unique fingerprint.

Cite this