arrow
Return

On text-based mining with active learning and background knowledge using SVM

delete2006-04-20
delete23
delete
OA
AI
C
Catarina Silva *
B
Bernardete Ribeiro
DOI:10.1007/s00500-006-0080-8delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Text mining, intelligent text analysis, text data mining and knowledge-discovery in text are generally used aliases to the process of extracting relevant and non-trivial information from text. Some crucial issues arise when trying to solve this problem, such as document representation and deficit of labeled data. This paper addresses these problems by introducing information from unlabeled documents in the training set, using the support vector machine (SVM) separating margin as the differentiating factor. Besides studying the influence of several pre-processing methods and concluding on their relative significance, we also evaluate the benefits of introducing background knowledge in a SVM text classifier. We further evaluate the possibility of actively learning and propose a method for successfully combining background knowledge and active learning. Experimental results show that the proposed techniques, when used alone or combined, present a considerable improvement in classification performance, even when small labeled training sets are available.
Keywords:
text mining
partially labeled data
support vector machines

Journal

Soft Computing cover
Soft Computing
IF:
2.5
Papers:
1.0W
Citations:
2.1W

Organization

No organization information available
Cited Papers

Cited Papers

A dynamic programming approach to missing data estimation using neural networks
err2013-07-01
err0
PREAI
errFulufhelo V. Nelwamondo; Dan Golding; Tshilidzi Marwala
errShare
errSave