arrow
返回

Transformer-BLS: An efficient learning algorithm based on multi-head attention mechanism and incremental learning algorithms

delete2024-03-01
delete10
PRE
AI
付
付荣荣 (Rongrong Fu)
H
Haifeng Liang
S
Shiwei Wang
C
Chengcheng Jia
T
Tengfei Gao
陈
陈丹 (Dan Chen)
Y
Yaodong Wang *
DOI:10.1016/j.eswa.2023.121734delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Due to its efficient model calibration given by unique incremental learning capability, broad learning system (BLS) has made impressive progress in image analytical tasks such as image classification and object detection. Inspired by this incremental remodel success, we proposed a novel transformer-BLS network to achieve a trade-off between model training speed and accuracy. Specially, we developed sub-BLS layers with the multi-head attention mechanism and combining these layers to construct a transformer-BLS network. In particular, our proposed transformer-BLS network provides four different incremental learning algorithms that enable the proposed model can realize the increments of its feature nodes, enhancement nodes, input data and sub-BLS layers, respectively, without the need of the full-weight update in this model. Furthermore, we validated the performance of our transformer-BLS network and its four incremental learning algorithms on a variety of image classification datasets. The results demonstrated that the proposed transformer-BLS maintains classification performance on both the MNIST and Fashion-MNIST datasets, while saving 2/3 of the training time. These findings imply that the proposed method has the potential in significant reducing model training complexity with this incremental remodel system, while simultaneously improving the increment learning performance of the original BLS within such contexts, especially in the classification task of some datasets.
Keyword:
Broad learning system
Multi-head attention mechanism
Transformer structure
Incremental learning algorithms

期刊

Expert Systems with Applications 封面图
Expert Systems with Applications
IF:
7.5
论文数:
3.0W
被引数:
10.2W

机构

T
Toronto Metropolitan University
学者数:
6.0K
论文数: 7.0K
被引数: 6.4K
Y
Yanshan University
学者数:
1.7W
论文数: 1.1W
被引数: 1.3W
T
technology & engineering center for space utilization, cas
学者数:
136
论文数: 116
被引数: 0
W
wuhan university
学者数:
8.1W
论文数: 5.8W
被引数: 70
C
chinese academy of sciences
学者数:
56.7W
论文数: 45.0W
被引数: 704
学者 查看更多机构
引用论文

引用论文

A defense method against backdoor attacks on neural networks
err2023-03-01
err10
PREAI
errKaviani, Sara; Shamshiri, Samaneh; Sohn, Insoo
err分享
err收藏
Stacked Broad Learning System: From Incremental Flatted Structure to Deep Model
err2021-01-01
err86
PREAI
errLiu, Zhulin; Chen, C. L. Philip; Feng, Shuang; Feng, Qiying; Zhang, Tong
err分享
err收藏
From WASD to BLS with application to pattern classification
err2021-09-01
err11
PREAI
errLiu, Mei; Li, Hongwei; Li, Yan; Jin, Long; Huang, Zhiguan
err分享
err收藏
学者 查看更多内容