arrow
Return

Polyseme-Aware Vector Representation for Text Classification

delete2020-01-01
delete1
delete
OA
AI
S
Shun Guo
N
Nianmin Yao *
DOI:10.1109/ACCESS.2020.3010981delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
Representation models for text classification have recently shown impressive performance. However, these models neglect the importance of polysemous words in text. When polysemous words appear in a text, imprecise polysemous word embeddings will produce low-quality text representation that results in changing the original meaning of the text. To address this problem, in this paper, we present a more effective model architecture, the polyseme-aware vector representation model (PAVRM), to generate more precise vector representations for words and texts. The PAVRM can effectively identify polysemous words in a corpus with a context clustering algorithm. Additionally, we propose two methods to construct polysemous word representations, PAVRM-Context and PAVRM-Center. Experiments conducted on three standard text classification tasks and a custom text classification task demonstrate that the proposed PAVRM can be effectively introduced into existing models to generate higher-quality word and text representations to achieve better classification performance.
Keywords:
Task analysis
Semantics
Text categorization
Training
Computational modeling
Context modeling
Microsoft Windows
Polysemous words
context clustering algorithm
PAVRM-Context
PAVRM-Center
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

IEEE Access cover
IEEE Access
IF:
3.6
Papers:
9.8W
Citations:
29.4W

Organization

D
Dalian University of Technology
Scholars:
5.9W
Papers: 4.4W
Citations: 5.5W