arrow
返回

Toward Deep Adaptive Hinging Hyperplanes

delete2022-11-01
delete5
delete
OA
AI
陶清华 封面图
陶清华 (Qinghua Tao)
J
Jun Xu
Z
Zhen Li
N
Na Xie *
S
Shuning Wang
李晓理 封面图
李晓理 (Xiaoli Li)
J
Johan A. K. Suykens
DOI:10.1109/TNNLS.2021.3079113delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
The adaptive hinging hyperplane (AHH) model is a popular piecewise linear representation with a generalized tree structure and has been successfully applied in dynamic system identification. In this article, we aim to construct the deep AHH (DAHH) model to extend and generalize the networking of AHH model for high-dimensional problems. The network structure of DAHH is determined through a forward growth, in which the activity ratio is introduced to select effective neurons and no connecting weights are involved between the layers. Then, all neurons in the DAHH network can be flexibly connected to the output in a skip-layer format, and only the corresponding weights are the parameters to optimize. With such a network framework, the backpropagation algorithm can be implemented in DAHH to efficiently tackle large-scale problems and the gradient vanishing problem is not encountered in the training of DAHH. In fact, the optimization problem of DAHH can maintain convexity with convex loss in the output layer, which brings natural advantages in optimization. Different from the existing neural networks, DAHH is easier to interpret, where neurons are connected sparsely and analysis of variance (ANOVA) decomposition can be applied, facilitating to revealing the interactions between variables. A theoretical analysis toward universal approximation ability and explicit domain partitions are also derived. Numerical experiments verify the effectiveness of the proposed DAHH.
Keyword:
Neurons
Artificial neural networks
Network topology
Training
Topology
Optimization
Adaptive systems
Adaptive hinging hyperplanes (AHHs)
analysis of variance (ANOVA) decomposition
domain partition
piecewise linear (PWL)
skip-layer connection
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

IEEE Transactions on Neural Networks and Learning Systems 封面图
IEEE Transactions on Neural Networks and Learning Systems
IF:
8.9
论文数:
7.5K
被引数:
7.2W

机构

H
harbin institute of technology
学者数:
8.0W
论文数: 6.6W
被引数: 66
T
tsinghua university
学者数:
11.9W
论文数: 10.0W
被引数: 137
K
KU Leuven
学者数:
5.7W
论文数: 5.2W
被引数: 8.1W
B
Beijing University of Technology
学者数:
2.8W
论文数: 2.1W
被引数: 2.7W
C
central university of finance & economics
学者数:
1.8K
论文数: 2.0K
被引数: 2
学者 查看更多机构