arrow
返回

Enhancing Visual Coding Through Collaborative Perception

delete2023-12-01
delete1
delete
OA
AI
L
Lingling An
Z
Zhen Yan
W
Weizheng Wang *
J
Jian K. Liu
K
Keping Yu
DOI:10.1109/TCDS.2022.3203422delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
A central challenge facing the nature human-computer interaction involves understanding how neural circuits process visual perceptual information to improve the user's operation ability under complex tasks. Visual coding models aim to explore the biological characteristics of retinal ganglion cells to provide quantitative predictions of responses to a range of visual stimuli. The existing visual coding models lack adaptability in natural and complex scenes. Therefore this article proposes an enhanced visual coding model through collaborative perception. Our model first extracts the multimodal spatiotemporal features of the input video to simulate the retinal response characteristics adaptively. Second, it uses the basis function to compile the input stimulus into a multimodal stimulus matrix. Afterward, the upstream and downstream filters reform the stimulus matrix to generate the spike sequence. Experiments show that the proposed model reproduces the physiological characteristics of ganglion cells in the biological retina, leading to the high accuracy, good adaptability, and biological interpretability in comparison with its rivals.
Keyword:
Feature compilation
multimodal stimulus
nonlinearity
visual coding

期刊

IEEE Transactions on Cognitive and Developmental Systems 封面图
IEEE Transactions on Cognitive and Developmental Systems
IF:
4.9
论文数:
1.0K
被引数:
3.5K

机构

C
City University of Hong Kong
学者数:
2.3W
论文数: 3.0W
被引数: 6.1W
U
university of leeds
学者数:
3.6W
论文数: 3.3W
被引数: 45
H
Hosei University
学者数:
953
论文数: 1.1K
被引数: 718
X
Xidian University
学者数:
2.4W
论文数: 1.9W
被引数: 9.7K
学者 查看更多机构