Explaining and Interpreting LSTMs

Leila Arras; Jose A. Arjona-Medina; Michael Widrich; Gr\'egoire; Montavon; Michael Gillhofer; Klaus-Robert M\"uller; Sepp Hochreiter and; Wojciech Samek

arXiv:1909.12114·cs.LG·September 27, 2019

Explaining and Interpreting LSTMs

Leila Arras, Jose A. Arjona-Medina, Michael Widrich, Gr\'egoire, Montavon, Michael Gillhofer, Klaus-Robert M\"uller, Sepp Hochreiter and, Wojciech Samek

PDF

TL;DR

This paper adapts the Layer-wise Relevance Propagation technique to LSTM networks, providing a new method for explaining their predictions by accounting for LSTM-specific structures and interactions.

Contribution

It introduces a novel propagation scheme and theoretical extension to explain LSTM predictions, bridging a gap in interpretability methods for sequential models.

Findings

01

Effective LRP extension for LSTMs

02

Faithful explanations of LSTM predictions

03

Enhanced interpretability of sequential models

Abstract

While neural networks have acted as a strong unifying force in the design of modern AI systems, the neural network architectures themselves remain highly heterogeneous due to the variety of tasks to be solved. In this chapter, we explore how to adapt the Layer-wise Relevance Propagation (LRP) technique used for explaining the predictions of feed-forward networks to the LSTM architecture used for sequential data modeling and forecasting. The special accumulators and gated interactions present in the LSTM require both a new propagation scheme and an extension of the underlying theoretical framework to deliver faithful explanations.

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

MethodsSigmoid Activation · Tanh Activation · Long Short-Term Memory