Noisin: Unbiased Regularization for Recurrent Neural Networks

Dieng, Adji B.; Ranganath, Rajesh; Altosaar, Jaan; Blei, David M.

Statistics > Machine Learning

arXiv:1805.01500 (stat)

[Submitted on 3 May 2018 (v1), last revised 13 Jul 2018 (this version, v2)]

Title:Noisin: Unbiased Regularization for Recurrent Neural Networks

Authors:Adji B. Dieng, Rajesh Ranganath, Jaan Altosaar, David M. Blei

View PDF

Abstract:Recurrent neural networks (RNNs) are powerful models of sequential data. They have been successfully used in domains such as text and speech. However, RNNs are susceptible to overfitting; regularization is important. In this paper we develop Noisin, a new method for regularizing RNNs. Noisin injects random noise into the hidden states of the RNN and then maximizes the corresponding marginal likelihood of the data. We show how Noisin applies to any RNN and we study many different types of noise. Noisin is unbiased--it preserves the underlying RNN on average. We characterize how Noisin regularizes its RNN both theoretically and empirically. On language modeling benchmarks, Noisin improves over dropout by as much as 12.2% on the Penn Treebank and 9.4% on the Wikitext-2 dataset. We also compared the state-of-the-art language model of Yang et al. 2017, both with and without Noisin. On the Penn Treebank, the method with Noisin more quickly reaches state-of-the-art performance.

Comments:	In Proceedings of the International Conference on Machine Learning, 2018
Subjects:	Machine Learning (stat.ML); Machine Learning (cs.LG); Methodology (stat.ME)
Cite as:	arXiv:1805.01500 [stat.ML]
	(or arXiv:1805.01500v2 [stat.ML] for this version)
	https://doi.org/10.48550/arXiv.1805.01500

Submission history

From: Adji Bousso Dieng [view email]
[v1] Thu, 3 May 2018 18:34:52 UTC (247 KB)
[v2] Fri, 13 Jul 2018 00:08:47 UTC (238 KB)

Statistics > Machine Learning

Title:Noisin: Unbiased Regularization for Recurrent Neural Networks

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Statistics > Machine Learning

Title:Noisin: Unbiased Regularization for Recurrent Neural Networks

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators