Neural Response Interpretation through the Lens of Critical Pathways

Ashkan Khakzar; Soroosh Baselizadeh; Saurabh Khanduja; Christian; Rupprecht; Seong Tae Kim; Nassir Navab

arXiv:2103.16886·cs.CV·April 1, 2021

Neural Response Interpretation through the Lens of Critical Pathways

Ashkan Khakzar, Soroosh Baselizadeh, Saurabh Khanduja, Christian, Rupprecht, Seong Tae Kim, Nassir Navab

PDF

2 Repos

TL;DR

This paper investigates how to identify and utilize critical neural pathways for interpreting network responses, proposing a neuron contribution-based pathway selection method and a new feature attribution technique called 'pathway gradient' validated through experiments.

Contribution

It introduces a neuron contribution-based pathway selection method that ensures critical input features are included, and proposes 'pathway gradient' for feature attribution, validated by experiments.

Findings

01

Pathways selected via neuron contribution are locally linear.

02

Pathway gradient effectively attributes critical input features.

03

Selected pathways correspond to critical input features.

Abstract

Is critical input information encoded in specific sparse pathways within the neural network? In this work, we discuss the problem of identifying these critical pathways and subsequently leverage them for interpreting the network's response to an input. The pruning objective -- selecting the smallest group of neurons for which the response remains equivalent to the original network -- has been previously proposed for identifying critical pathways. We demonstrate that sparse pathways derived from pruning do not necessarily encode critical input information. To ensure sparse pathways include critical fragments of the encoded input information, we propose pathway selection via neurons' contribution to the response. We proceed to explain how critical pathways can reveal critical input features. We prove that pathways selected via neuron contribution are locally linear (in an L2-ball), a…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

MethodsPruning