arrow
Return

A document clustering algorithm for discovering and describing topics

delete2010-04-01
delete27
PRE
AI
H
Henry Anaya-Sánchez *
R
Rafael Berlanga
DOI:10.1016/j.patrec.2009.11.013delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
In this paper, we introduce a new clustering algorithm for discovering and describing the topics comprised in a text collection. Our proposal relies on both the most probable term pairs generated from the collection and the estimation of the topic homogeneity associated to these pairs Topics and their descriptions are generated from those term pairs whose support sets are homogeneous enough for representing collection topics Experimental results obtained over three benchmark text collections demonstrate the effectiveness and utility of this new approach (C) 2009 Published by Elsevier B V
Keywords:
Document clustering
Topic discovery
Topic description

Journal

Pattern Recognition Letters cover
Pattern Recognition Letters
IF:
3.3
Papers:
7.9K
Citations:
1.6W

Organization

U
Universitat Jaume I
Scholars:
4.7K
Papers: 4.8K
Citations: 6.1K
U
universidad de oriente santiago de cuba
Scholars:
296
Papers: 202
Citations: 0