Layer-Wise Interpretation of Deep Neural Networks Using Identity   Initialization

Shohei Kubota; Hideaki Hayashi; Tomohiro Hayase; Seiichi Uchida

arXiv:2102.13333·cs.LG·March 1, 2021

Layer-Wise Interpretation of Deep Neural Networks Using Identity Initialization

Shohei Kubota, Hideaki Hayashi, Tomohiro Hayase, Seiichi Uchida

PDF

Open Access

TL;DR

This paper introduces a novel interpretability method for deep neural networks using identity initialization, enabling layer-wise analysis of neuron contributions and roles in classification, which enhances transparency.

Contribution

The proposed method leverages identity initialization to analyze neuron contributions at each layer, revealing the roles of feature extraction and classification in deep networks.

Findings

01

Identity-initialized networks maintain near-identity weights after training.

02

Layer-wise contribution maps can be generated to visualize neuron impact.

03

The method allows recognition accuracy assessment at each layer.

Abstract

The interpretability of neural networks (NNs) is a challenging but essential topic for transparency in the decision-making process using machine learning. One of the reasons for the lack of interpretability is random weight initialization, where the input is randomly embedded into a different feature space in each layer. In this paper, we propose an interpretation method for a deep multilayer perceptron, which is the most general architecture of NNs, based on identity initialization (namely, initialization using identity matrices). The proposed method allows us to analyze the contribution of each neuron to classification and class likelihood in each hidden layer. As a property of the identity-initialized perceptron, the weight matrices remain near the identity matrices even after learning. This property enables us to treat the change of features from the input to each hidden layer as…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsAdversarial Robustness in Machine Learning · Explainable Artificial Intelligence (XAI) · Advanced Neural Network Applications