arrow
Return

Improving a statistical language model through non-linear prediction

delete2009-03-01
delete12
PRE
AI
A
Andriy Mnih *
Y
Yuecheng Zhang
G
Geoffrey E. Hinton
DOI:10.1016/j.neucom.2008.12.025delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
We show how to improve a state-of-the-art neural network language model that converts the previous context words into feature vectors and combines these feature vectors linearly to predict the feature vector of the next word. Significant improvements in predictive accuracy are achieved by using a non-linear subnetwork to modulate the effects of the context words or to produce a non-linear correction term when predicting the feature vector. A log-bilinear language model that incorporates both of these improvements achieves a 26% reduction in perplexity over the best n-gram model on a fairly large dataset. (C) 2009 Elsevier B.V. All rights reserved.
Keywords:
Statistical language modelling
Distributed representations
Neural networks
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Neurocomputing cover
Neurocomputing
IF:
6.5
Papers:
2.5W
Citations:
6.5W

Organization

U
university of toronto
Scholars:
14.7W
Papers: 12.0W
Citations: 165