arrow
Return

Multi-Document Summarization Using Selective Attention Span and Reinforcement Learning

delete2023-01-01
delete1
PRE
AI
Y
Yash Kumar Atri *
V
Vikram Goyal
T
Tanmoy Chakraborty
DOI:10.1109/TASLP.2023.3316459delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
The Abstractive text summarization systems using recently improved RNN-based sequence-to-sequence architecture have shown great promise for single-document summarization. However, such neural models fail to perpetuate the performance in the multi-document summarization setting owing to the long-range dependencies within the documents, overlapping/contradicting facts and extrinsic model hallucinations. These shortcomings augment the model to generate inconsistent, repetitive and non-factual summaries. In this work, we introduce REISA, a sequence-to-sequence model with a novel reinforced selective attention span that attends over the input and recalibrates the local attention weights to focus on important segments while generating output at each time step. REISA utilizes a reinforcement learning-based policy gradient algorithm to reward the model and formulate attention distributions over the encoder input. We further benchmark REISA on two widely-used multi-document summarization corpora - Multinews and CQASumm, and observe an improvement of +2.91 and +6.64 ROUGE-L scores, respectively. The qualitative analyses on semantic similarity by BERTScore, faithfulness by question-answer evaluation and human evaluation show significant improvement over the baseline-generated summaries.
Keywords:
Abstractive Text Summarization
Multi-Document Summarization
seq2seq
Deep Reinforcement Learning

Journal

I
IEEE-ACM Transactions on Audio Speech and Language Processing
IF:
5.1
Papers:
2.6K
Citations:
1.1W

Organization

I
indian institute of technology system (iit system)
Scholars:
9.5W
Papers: 9.9W
Citations: 93
I
Indraprastha Institute of Information Technology Delhi
Scholars:
933
Papers: 689
Citations: 558
Cited Papers

Cited Papers

SCALING WITH KNOWN UNCERTAINTY: A SYNTHESIS
err2006-01-01
err0
PREAI
errJIANGUO WU; HARBIN LI; K. BRUCE JONES; ORIE L. LOUCKS
errShare
errSave
Successful treatment with alectinib after crizotinib-induced interstitial lung disease
err2015-12-01
err0
PREAI
errHaruka Chino; Akimasa Sekine; Hideya Kitamura; Terufumi Kato; Takashi Ogura
errShare
errSave
researcher View more