arrow
Return

Stimulus-driven and concept-driven analysis for image caption generation

delete2020-07-01
delete168
PRE
AI
S
Songtao Ding
S
Shiru Qu
Y
Yuling Xi
S
Shaohua Wan *
DOI:10.1016/j.neucom.2019.04.095delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Recently, image captioning has achieved great progress in computer vision and artificial intelligence. However, language models still failed to achieve the desired results in high-level visual tasks. Generating accurate image captions for a complex scene that contains multiple targets is a challenge. To solve these problems, we introduce the theory of attention in psychology to image caption generation. We propose two types of attention mechanisms: The stimulus-driven and the concept-driven. Our attention model relies on a combination of convolutional neural network (CNN) over images and long-short term memory (LSTM) network over sentences. Comparison of experimental results illustrates that our proposed method achieves good performance on the MSCOCO test server. (C) 2019 Elsevier B.V. All rights reserved.
Keywords:
Image captioning
Stimulus-driven
Concept-driven
Attention mechanism
LSTM
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Neurocomputing cover
Neurocomputing
IF:
6.5
Papers:
2.5W
Citations:
6.5W

Organization

Z
zhongnan university of economics & law
Scholars:
2.0K
Papers: 2.2K
Citations: 3
N
Northwestern Polytechnical University
Scholars:
4.6W
Papers: 3.7W
Citations: 5.3W