An Analysis of Attention over Clinical Notes for Predictive Tasks

Sarthak Jain; Ramin Mohammadi; Byron C. Wallace

arXiv:1904.03244·cs.CL·April 9, 2019

An Analysis of Attention over Clinical Notes for Predictive Tasks

Sarthak Jain, Ramin Mohammadi, Byron C. Wallace

PDF

TL;DR

This paper investigates the use of attention mechanisms in neural models analyzing clinical notes within electronic medical records, aiming to improve predictive performance and interpretability, but finds mixed results regarding their explanatory value.

Contribution

It demonstrates that attention mechanisms are essential for neural models to achieve competitive performance on EMR tasks, but their interpretability benefits are uncertain.

Findings

01

Attention improves neural model performance on EMR tasks.

02

Interpretability of attention in clinical notes remains unclear.

03

Attention mechanisms are critical for effective neural encoding of notes.

Abstract

The shift to electronic medical records (EMRs) has engendered research into machine learning and natural language technologies to analyze patient records, and to predict from these clinical outcomes of interest. Two observations motivate our aims here. First, unstructured notes contained within EMR often contain key information, and hence should be exploited by models. Second, while strong predictive performance is important, interpretability of models is perhaps equally so for applications in this domain. Together, these points suggest that neural models for EMR may benefit from incorporation of attention over notes, which one may hope will both yield performance gains and afford transparency in predictions. In this work we perform experiments to explore this question using two EMR corpora and four different predictive tasks, that: (i) inclusion of attention mechanisms is critical for…

Figures40

Click any figure to enlarge with its caption.

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

MethodsInterpretability