Subregular Complexity and Deep Learning

Avcu, Enes; Shibata, Chihiro; Heinz, Jeffrey

Computer Science > Computation and Language

arXiv:1705.05940 (cs)

[Submitted on 16 May 2017 (v1), last revised 14 Oct 2017 (this version, v3)]

Title:Subregular Complexity and Deep Learning

Authors:Enes Avcu, Chihiro Shibata, Jeffrey Heinz

View PDF

Abstract:This paper argues that the judicial use of formal language theory and grammatical inference are invaluable tools in understanding how deep neural networks can and cannot represent and learn long-term dependencies in temporal sequences. Learning experiments were conducted with two types of Recurrent Neural Networks (RNNs) on six formal languages drawn from the Strictly Local (SL) and Strictly Piecewise (SP) classes. The networks were Simple RNNs (s-RNNs) and Long Short-Term Memory RNNs (LSTMs) of varying sizes. The SL and SP classes are among the simplest in a mathematically well-understood hierarchy of subregular classes. They encode local and long-term dependencies, respectively. The grammatical inference algorithm Regular Positive and Negative Inference (RPNI) provided a baseline. According to earlier research, the LSTM architecture should be capable of learning long-term dependencies and should outperform s-RNNs. The results of these experiments challenge this narrative. First, the LSTMs' performance was generally worse in the SP experiments than in the SL ones. Second, the s-RNNs out-performed the LSTMs on the most complex SP experiment and performed comparably to them on the others.

Subjects:	Computation and Language (cs.CL)
Cite as:	arXiv:1705.05940 [cs.CL]
	(or arXiv:1705.05940v3 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.1705.05940

Submission history

From: Enes Avcu [view email]
[v1] Tue, 16 May 2017 22:13:45 UTC (4,057 KB)
[v2] Mon, 3 Jul 2017 02:18:14 UTC (4,057 KB)
[v3] Sat, 14 Oct 2017 18:24:08 UTC (76 KB)

Computer Science > Computation and Language

Title:Subregular Complexity and Deep Learning

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Subregular Complexity and Deep Learning

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators