Discovery of Natural Language Concepts in Individual Units of CNNs

Seil Na; Yo Joong Choe; Dong-Hyun Lee; Gunhee Kim

arXiv:1902.07249·cs.CL·March 1, 2019·5 cites

Discovery of Natural Language Concepts in Individual Units of CNNs

Seil Na, Yo Joong Choe, Dong-Hyun Lee, Gunhee Kim

PDF

Open Access 1 Repo

TL;DR

This paper investigates how individual units in CNNs trained on language tasks respond to specific linguistic concepts, revealing that units are selectively responsive to morphemes, words, and phrases, thus shedding light on the internal representations of deep language models.

Contribution

The paper introduces a concept alignment method to analyze unit responses and demonstrates that CNN units are selectively responsive to linguistic concepts across various architectures and tasks.

Findings

01

Units respond to specific morphemes, words, and phrases.

02

Analysis across multiple models and datasets confirms selective responsiveness.

03

Provides new insights into how CNNs understand natural language.

Abstract

Although deep convolutional networks have achieved improved performance in many natural language tasks, they have been treated as black boxes because they are difficult to interpret. Especially, little is known about how they represent language in their intermediate layers. In an attempt to understand the representations of deep convolutional networks trained on language tasks, we show that individual units are selectively responsive to specific morphemes, words, and phrases, rather than responding to arbitrary and uninterpretable patterns. In order to quantitatively analyze such an intriguing phenomenon, we propose a concept alignment method based on how units respond to the replicated text. We conduct analyses with different architectures on multiple datasets for classification and translation tasks and provide new insights into how deep models understand natural language.

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

seilna/CNN-Units-in-NLP
tfOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsTopic Modeling · Natural Language Processing Techniques · Text and Document Classification Technologies