arrow
Return

Active Vision for Deep Visual Learning: A Unified Pooling Framework

delete2022-10-01
delete3
delete
OA
AI
N
Nan Guo
K
Ke Gu
J
Junfei Qiao *
H
Hantao Liu
DOI:10.1109/TII.2021.3129813delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Convolutional neural networks (CNNs) can be generally regarded as learning-based visual systems for computer vision tasks. By imitating the operating mechanism of the human visual system (HVS), CNNs can even achieve better results than human beings in some visual tasks. However, they are primary when compared to the HVS for the reason that the HVS has the ability of active vision to promptly analyze and adapt to specific tasks. In this article, a new unified pooling framework is proposed and a series of pooling methods are designed based on the framework to implement active vision to CNNs. In addition, an active selection pooling (ASP) is put forward to reorganize the existing and newly proposed pooling methods. The CNN models with an ASP tend to have a behavior of focus selection according to tasks during the training process, which acts extremely similar to the HVS.
Keywords:
Visual systems
Task analysis
Visualization
Training
Convolutional neural networks
Informatics
Image color analysis
Active vision
deep convolutional neural networks (CNNs)
deep visual learning
human visual system (HVS)
pooling framework

Journal

IEEE Transactions on Industrial Informatics cover
IEEE Transactions on Industrial Informatics
IF:
9.9
Papers:
8.3K
Citations:
6.0W

Organization

C
Cardiff University
Scholars:
2.7W
Papers: 2.5W
Citations: 3.5W
B
Beijing University of Technology
Scholars:
2.8W
Papers: 2.1W
Citations: 2.7W