返回
Efficient English text classification using selected Machine Learning Techniques
DOI:10.1016/j.aej.2021.02.009.png)
摘要
En 中文
Text classification (TC) is an approach used for the classification of any kind of documents for the target category or out. In this paper, we implemented the Support Vector Machines (SVM) model in classifying English text and documents. Here we did two analytical experiments to check the selected classifiers using English documents. Experimental results performed on a set of 1033 text document present that the Rocchio classifier provides the best performance results when the size of the feature set is small while SVM outperforms the other classifiers. From the experimental analysis, we observed that the classification rate exceeds 90% when using more than 4000 features. (C) 2021 THE AUTHOR. Published by Elsevier BV on behalf of Faculty of Engineering, Alexandria University. This is an open access article under the CC BY-NC-ND license (http://creativecommons.org/ licenses/by-nc-nd/4.0/).
Keyword:
Text classification
English language
Machine Learning
Text mining
Support Vector Machines
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
6.8
论文数:
6.3K
被引数:
2.6W
机构
引用论文
Transformation of Summary Statistics from Linear Mixed Model Association on All-or-None Traits to Odds Ratio
GENETICS
IF5.1
Using text mining and sentiment analysis for online forums hotspot detection and forecast使用文本挖掘和情感分析进行在线论坛热点检测和预测
High dimensional data classification and feature selection using support vector machines基于支持向量机的高维数据分类与特征选择
Nonobese diabetic/severe combined immunodeficient murine xenograft model for human uterine leiomyoma

