Discovery of Natural Language Concepts in Individual Units of CNNs

Na, Seil; Choe, Yo Joong; Lee, Dong-Hyun; Kim, Gunhee

Computer Science > Computation and Language

arXiv:1902.07249 (cs)

[Submitted on 18 Feb 2019 (v1), last revised 28 Feb 2019 (this version, v2)]

Title:Discovery of Natural Language Concepts in Individual Units of CNNs

Authors:Seil Na, Yo Joong Choe, Dong-Hyun Lee, Gunhee Kim

View PDF

Abstract:Although deep convolutional networks have achieved improved performance in many natural language tasks, they have been treated as black boxes because they are difficult to interpret. Especially, little is known about how they represent language in their intermediate layers. In an attempt to understand the representations of deep convolutional networks trained on language tasks, we show that individual units are selectively responsive to specific morphemes, words, and phrases, rather than responding to arbitrary and uninterpretable patterns. In order to quantitatively analyze such an intriguing phenomenon, we propose a concept alignment method based on how units respond to the replicated text. We conduct analyses with different architectures on multiple datasets for classification and translation tasks and provide new insights into how deep models understand natural language.

Comments:	Published as a conference paper at ICLR 2019
Subjects:	Computation and Language (cs.CL); Machine Learning (cs.LG); Machine Learning (stat.ML)
Cite as:	arXiv:1902.07249 [cs.CL]
	(or arXiv:1902.07249v2 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.1902.07249

Submission history

From: Seil Na [view email]
[v1] Mon, 18 Feb 2019 06:19:14 UTC (2,279 KB)
[v2] Thu, 28 Feb 2019 09:05:10 UTC (2,277 KB)

Full-text links:

Access Paper:

view license

Current browse context:

stat

< prev | next >

new | recent | 2019-02

Change to browse by:

cs
cs.CL
cs.LG
stat.ML

References & Citations

DBLP - CS Bibliography

listing | bibtex

Seil Na
Yo Joong Choe
Dong-Hyun Lee
Gunhee Kim

export BibTeX citation

Computer Science > Computation and Language

Title:Discovery of Natural Language Concepts in Individual Units of CNNs

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Discovery of Natural Language Concepts in Individual Units of CNNs

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators