Estimating marginal probabilities of n-grams for recurrent neural language models

Thanapon Noraset; Doug Downey; Lidong Bing

Conference ProceedingsOPEN ACCESS

Estimating marginal probabilities of n-grams for recurrent neural language models

Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing, EMNLP 2018 (2018) 2930-2935

DOI: 10.18653/v1/d18-1322

2Citations

88Readers

Abstract

Recurrent neural network language models (RNNLMs) are the current standard-bearer for statistical language modeling. However, RNNLMs only estimate probabilities for complete sequences of text, whereas some applications require context-independent phrase probabilities instead. In this paper, we study how to compute an RNNLM's marginal probability: the probability that the model assigns to a short sequence of text when the preceding context is not known. We introduce a simple method of altering the RNNLM training to make the model more accurate at marginal estimation. Our experiments demonstrate that the technique is effective compared to baselines including the traditional RNNLM probability and an importance sampling approach. Finally, we show how we can use the marginal estimation to improve an RNNLM by training the marginals to match n-gram probabilities from a larger corpus.

Cite

CITATION STYLE

APA

Noraset, T., Downey, D., & Bing, L. (2018). Estimating marginal probabilities of n-grams for recurrent neural language models. In Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing, EMNLP 2018 (pp. 2930–2935). Association for Computational Linguistics. https://doi.org/10.18653/v1/d18-1322

Estimating marginal probabilities of n-grams for recurrent neural language models

Abstract

Cite

Register to see more suggestions