arrow
Return

DeepFix: A Fully Convolutional Neural Network for Predicting Human Eye Fixations

delete2017-09-01
delete323
delete
OA
AI
S
Srinivas S S Kruthiventi *
K
Kumar Ayush
R
R. Venkatesh Babu
DOI:10.1109/TIP.2017.2710620delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
Understanding and predicting the human visual attention mechanism is an active area of research in the fields of neuroscience and computer vision. In this paper, we propose DeepFix, a fully convolutional neural network, which models the bottom-up mechanism of visual attention via saliency prediction. Unlike classical works, which characterize the saliency map using various hand-crafted features, our model automatically learns features in a hierarchical fashion and predicts the saliency map in an end-to-end manner. DeepFix is designed to capture semantics at multiple scales while taking global context into account, by using network layers with very large receptive fields. Generally, fully convolutional nets are spatially invariant-this prevents them from modeling location-dependent patterns (e.g., centre-bias). Our network handles this by incorporating a novel location-biased convolutional layer. We evaluate our model on multiple challenging saliency data sets and show that it achieves the state-of-the-art results.
Keywords:
Saliency prediction
eye fixations
convolutional neural network
deep learning
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

IEEE Transactions on Image Processing cover
IEEE Transactions on Image Processing
IF:
13.7
Papers:
1.0W
Citations:
8.4W

Organization

I
indian institute of technology system (iit system)
Scholars:
9.5W
Papers: 9.9W
Citations: 93
I
indian institute of science (iisc) - bangalore
Scholars:
1.4W
Papers: 1.4W
Citations: 11