返回
Cross-Lingual Adaptation Using Structural Correspondence Learning
DOI:10.1145/2036264.2036277.png)
摘要
En 中文
Cross-lingual adaptation is a special case of domain adaptation and refers to the transfer of classification knowledge between two languages. In this article we describe an extension of Structural Correspondence Learning (SCL), a recently proposed algorithm for domain adaptation, for cross-lingual adaptation in the context of text classification. The proposed method uses unlabeled documents from both languages, along with a word translation oracle, to induce a cross-lingual representation that enables the transfer of classification knowledge from the source to the target language. The main advantages of this method over existing methods are resource efficiency and task specificity. We conduct experiments in the area of cross-language topic and sentiment classification involving English as source language and German, French, and Japanese as target languages. The results show a significant improvement of the proposed method over a machine translation baseline, reducing the relative error due to cross-lingual adaptation by an average of 30% (topic classification) and 59% (sentiment classification). We further report on empirical analyses that reveal insights into the use of unlabeled data, the sensitivity with respect to important hyperparameters, and the nature of the induced cross-lingual word correspondences.
Keyword:
Algorithms
Experimentation
Performance
Cross-language text classification
cross-lingual adaptation
structural correspondence learning
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
6.6
论文数:
1.5K
被引数:
6.2K
机构
引用论文
Aberrant DNA Methylation Is Associated with Disease Progression, Resistance to Imatinib and Shortened Survival in Chronic Myelogenous Leukemia
PLoS ONE
IF0

