arrow
返回

Enhancing aspect-based sentiment analysis using data augmentation based on back-translation

delete2024-08-14
delete2
PRE
AI
A
Alireza Taheri
A
Azadeh Zamanifar *
A
Amirfarhad Farhadi
DOI:10.1007/s41060-024-00622-wdelete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Aspect-based sentiment analysis (ABSA) identifies mentioned aspects and predicts their associated sentiments in sentences. With the rapid growth of users' online activities, ABSA, as a means of automatically interpreting text, shows significant importance. Aspect terms and opinions are dissimilar for different topics; hence, providing more labeled data related to the domain might be required to achieve better performance. Labeling massive amounts of data is expensive and time-consuming, but by using data augmentation, enhancing the performance is possible without collecting new labeled data. In this paper, we present a hybrid data augmentation method to extend the original data and increase ABSA performance. Back-translation has proved helpful in other areas of NLP as a paraphrasing method to augment text, but because of the data structure of ABSA, it has not been used at its full potential in this field yet. We utilize Special Character Insertion (SCI) in back-translation to address this compatibility issue and generate synthetic augmented sentences. By doing so, the output sentences will preserve meaning and sentiments toward the aspect terms, and their location in the sentence will be restored. Using Random Entity Replacement (RER), we can make the back-translated sentence even more diverse to make the model generalize better on the limited data. RER is used to replace named entities with the lowest chance of getting replaced by the back-translator. Different ABSA models and two benchmark datasets, SemEval 2014 in Restaurant and Laptop domains, are used to evaluate our approach, and five different languages as the middle language for back-translation are investigated. Results show that using German as a middle language, our approach on average can increase accuracy and f1 score by 0.78 and 1.75 percent compared to the original dataset.
Keyword:
Aspect-based sentiment analysis
Data augmentation
Back-translation
Special character utilization

期刊

I
International Journal of Data Science and Analytics
IF:
2.8
论文数:
1.1K
被引数:
1.3K

机构

I
Islamic Azad University
学者数:
4.0W
论文数: 3.3W
被引数: 9.8K
引用论文

引用论文

Tailored text augmentation for sentiment analysis
err2022-11-01
err13
PREAI
errFeng, Zijian; Zhou, Hanzhang; Zhu, Zixiao; Mao, Kezhi
err分享
err收藏
PM2.5 emissions from different types of heavy-duty truck: a case study and meta-analysis of the Beijing-Tianjin-Hebei region
err2017-03-14
err0
PREAI
errLiying Song; Hongqing Song; Jingyi Lin; Cheng Wang; Mingxu Yu; Xiaoxia Huang; Yu Guan; Xing Wang; Li Du
err分享
err收藏
A Lexicon-Enhanced Attention Network for Aspect-Level Sentiment Analysis用于方面级情感分析的词典增强的注意力网络
err2020-01-01
err25
errOAAI
errRen, Zhiying; Zeng, Guangping; Chen, Liu; Zhang, Qingchuan; Zhang, Chunguang; Pan, Dingqi
err分享
err收藏
Hierarchical Data Augmentation and the Application in Text Classification
err2019-01-01
err19
errOAAI
errYu, Shujuan; Yang, Jie; Liu, Danlei; Li, Runqi; Zhang, Yun; Zhao, Shengmei
err分享
err收藏
Exogenous application of nitric oxide and spermidine reduces the negative effects of salt stress on tomato
err2017-12-15
err0
PREAI
errManzer H. Siddiqui; Saud A. Alamri; Mutahhar Y. Al-Khaishany; Mohammed A. Al-Qutami; Hayssam M. Ali; Hala AL-Rabiah; Hazem M. Kalaji
err分享
err收藏
Text data augmentations: Permutation, antonyms and negation
err2021-09-01
err21
errOAAI
errHaralabopoulos, Giannis; Torres, Mercedes Torres; Anagnostopoulos, Ioannis; McAuley, Derek
err分享
err收藏
A dependency syntactic knowledge augmented interactive architecture for end-to-end aspect-based sentiment analysis
err2021-09-01
err52
errOAAI
errLiang, Yunlong; Meng, Fandong; Zhang, Jinchao; Chen, Yufeng; Xu, Jinan; Zhou, Jie
err分享
err收藏
学者 查看更多内容