arrow
返回

A complex history browsing text categorization method with improved BERT embedding layer

delete2025-02-03
delete1
PRE
AI
Y
Yuanhang Wang
周
周永华 (Yonghua Zhou) *
H
Huiyu Qi
W
Wang, Dingyi
H
Huang, Annan
DOI:10.1007/s10489-025-06298-4delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
For long texts composed of multiple short fragments, the importance of each fragment to the classification task varies. Some fragments have higher discriminative power and positively contribute to the classification, while others lack discriminative power or even mislead it. Existing methods struggle to converge when handling texts with negative examples. This study analyzes user behavior and assigns interest scores to text fragments based on their classification relevance, allowing the model to focus more on important fragments. Building on bidirectional encoder representations from transformers (BERT), we propose an interest encoding layer model for historical browsing texts. By analyzing user behavior and incorporating an improved term frequency-inverse document frequency (TF-IDF) method, the model adds indicators to fragments with higher discriminative power for user behavior analysis, enabling the model to focus more on these during training. Finally, comparative experiments on the BERT model series validate the advantages of the proposed approach.
Keyword:
History browsing text
BERT network
TF-IDF method
Interest embedding
Text categorization

期刊

Applied Intelligence 封面图
Applied Intelligence
IF:
3.5
论文数:
7.6K
被引数:
1.7W

机构

B
Beijing Jiaotong University
学者数:
2.2W
论文数: 1.7W
被引数: 1.2W
引用论文

引用论文

err分享
err收藏
Deceptive reviews and sentiment polarity: Effective link by exploiting BERT
err2022-12-01
err13
PREAI
errCatelli, Rosario; Fujita, Hamido; De Pietro, Giuseppe; Esposito, Massimo
err分享
err收藏
学者 查看更多内容