arrow
返回

COSUM: Text summarization based on clustering and optimization

delete2018-10-12
delete69
delete
OA
AI
R
Rasim Alguliyev
R
Ramiz M. Aliguliyev *
N
Nijat R. Isazade
A
Asad Abdi
N
Norisma Idris
DOI:10.1111/exsy.12340delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
Text summarization is a process of extracting salient information from a source text and presenting that information to the user in a condensed form while preserving its main content. In the text summarization, most of the difficult problems are providing wide topic coverage and diversity in a summary. Research based on clustering, optimization, and evolutionary algorithm for text summarization has recently shown good results, making this a promising area. In this paper, for a text summarization, a two-stage sentences selection model based on clustering and optimization techniques, called COSUM, is proposed. At the first stage, to discover all topics in a text, the sentences set is clustered by using k-means method. At the second stage, for selection of salient sentences from clusters, an optimization model is proposed. This model optimizes an objective function that expressed as a harmonic mean of the objective functions enforcing the coverage and diversity of the selected sentences in the summary. To provide readability of a summary, this model also controls length of sentences selected in the candidate summary. For solving the optimization problem, an adaptive differential evolution algorithm with novel mutation strategy is developed. The method COSUM was compared with the 14 state-of-the-art methods: DPSO-EDASum; LexRank; CollabSum; UnifiedRank; 0-1 non-linear; query, cluster, summarize; support vector machine; fuzzy evolutionary optimization model; conditional random fields; MA-SingleDocSum; NetSum; manifold ranking; ESDS-GHS-GLO; and differential evolution, using ROUGE tool kit on the DUC2001 and DUC2002 data sets. Experimental results demonstrated that COSUM outperforms the state-of-the-art methods in terms of ROUGE-1 and ROUGE-2 measures.
Keyword:
adaptive differential evolution algorithm
content coverage
harmonic mean
information diversity
k-means
optimization model
sentence clustering
text summarization
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Expert Systems 封面图
Expert Systems
IF:
2.3
论文数:
2.6K
被引数:
3.8K

机构

R
RWTH Aachen University
学者数:
3.5W
论文数: 2.6W
被引数: 3.6W
A
azerbaijan national academy of sciences (anas)
学者数:
1.3K
论文数: 1.4K
被引数: 0
U
Universiti Teknologi Malaysia
学者数:
1.4W
论文数: 1.1W
被引数: 85
学者 查看更多机构
引用论文

引用论文

An unsupervised approach to generating generic summaries of documents
err2015-09-01
err31
PREAI
errAlguliyev, Rasim M.; Aliguliyev, Ramiz M.; Isazade, Nijat R.
err分享
err收藏
Incremental learning for ν-Support Vector RegressionΝ-支持向量回归的增量学习
err2015-07-01
err432
PREAI
errGu, Bin; Sheng, Victor S.; Wang, Zhijie; Ho, Derek; Osman, Said; Li, Shuo
err分享
err收藏
GenDocSum plus MCLR: Generic document summarization based on maximum coverage and less redundancy
err2012-11-01
err42
PREAI
errAlguliev, Rasim M.; Aliguliyev, Ramiz M.; Hajirahimova, Makrufa S.
err分享
err收藏
学者 查看更多内容