返回
Improving a statistical language model through non-linear prediction
DOI:10.1016/j.neucom.2008.12.025.png)
摘要
En 中文
We show how to improve a state-of-the-art neural network language model that converts the previous context words into feature vectors and combines these feature vectors linearly to predict the feature vector of the next word. Significant improvements in predictive accuracy are achieved by using a non-linear subnetwork to modulate the effects of the context words or to produce a non-linear correction term when predicting the feature vector. A log-bilinear language model that incorporates both of these improvements achieves a 26% reduction in perplexity over the best n-gram model on a fairly large dataset. (C) 2009 Elsevier B.V. All rights reserved.
Keyword:
Statistical language modelling
Distributed representations
Neural networks
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
6.5
论文数:
2.5W
被引数:
6.5W
机构
引用论文
没有更多内容

