返回
Bilingual recursive neural network based data selection for statistical machine translation
DOI:10.1016/j.knosys.2016.05.003.png)
摘要
En 中文
Data selection is a widely used and effective solution to domain adaptation in statistical machine translation (SMT). The dominant methods are perplexity-based ones, which do not consider the mutual translations of sentence pairs and tend to select short sentences. In this paper, to address these problems, we propose bilingual semi-supervised recursive neural network data selection methods to differentiate domain-relevant data from out-domain data. The proposed methods are evaluated in the task of building domain-adapted SMT systems. We present extensive comparisons and show that the proposed methods outperform the state-of-the-art data selection approaches. (C) 2016 Elsevier B.V. All rights reserved.
Keyword:
Data selection
Machine translation
Domain adaptation
Recursive neural network
Autoencoder
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
K
IF:
7.6
论文数:
1.3W
被引数:
4.5W
机构
引用论文
没有更多内容

