arrow
返回

Mood-aware visual question answering

delete2019-02-01
delete16
PRE
AI
N
Nelson Ruwa
Q
Qirong Mao *
L
Liangjun Wang
J
Jianping Gou
M
Ming Dong
DOI:10.1016/j.neucom.2018.11.049delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
The concept of Visual Question Answering (VQA) has recently attracted the attention of many researchers in the field of machine learning. Different attention models have been proposed in VQA for the purpose of addressing the need to focus on local regions of an image. This paper proposes the concept of Mood-Aware Visual Question Answering (MAVQA) using novel long short term memory (LSTM) and convolutional neural network (CNN) attention models that combine the local image features, the question and the mood detected from the particular region of the image to produce a mood-based answer using a pre-processed image dataset. The attention mechanisms serve to enable the VQA model to only focus on parts of the image that are relevant to both the detected mood and the key words in the question. The irrelevant parts of the image are ignored, thus improving classification accuracy by reducing the chances of predicting wrong answers. Whereas previous efforts have utilized CNN mostly for the embedding of images and text, we formulate a CNN attention algorithm for the image, question and mood. The more direct convolutional attention operation is more efficient and effective, when the number of views and kernel length are optimized, than the winding recurrent LSTM attention operation. The experimental results prove that MAVQA is effectively mood-aware, and the accuracy levels of our LSTM attention model are well within the range of previous conventional VQA benchmarks, while our novel CNN attention model outperforms the previous baselines in several instances. The additional attention on the mood does not only improve classification accuracy but also substantially contributes towards the analysis and comprehension of image features, a key development in modern artificial intelligence. (C) 2018 Published by Elsevier B.V.
Keyword:
Mood-aware
Visual question answering
Attention model
Long short term memory
Convolutional neural network
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Neurocomputing 封面图
Neurocomputing
IF:
6.5
论文数:
2.5W
被引数:
6.5W

机构

J
Jiangsu University
学者数:
4.0W
论文数: 2.8W
被引数: 5.5W
W
wayne state university
学者数:
2.0W
论文数: 1.6W
被引数: 17
引用论文

引用论文

Effects of label-dose permethrin administration in yearling beef cattle: I. Bull reproductive function and testicular histopathology
err2016-06-01
err0
PREAI
errTyler M. Dohlman; Patrick E. Phillips; Darin M. Madson; Christopher A. Clark; Patrick J. Gunn
err分享
err收藏
Uncovering the Temporal Context for Video Question Answering
err2017-07-13
err86
PREAI
errZhu, Linchao; Xu, Zhongwen; Yang, Yi; Hauptmann, Alexander G.
err分享
err收藏
Effects of pharmacological inhibition of dopamine receptors on memory load capacity
err2019-02-01
err0
PREAI
errLaura Olivito; Maria De Risi; Fabio Russo; Elvira De Leonibus
err分享
err收藏
Digital Twin Modeling Method for Hierarchical Stiffened Plate Based on Transfer Learning
err2023-01-09
err0
errOAAI
errZiyu Xu; Tianhe Gao; Zengcong Li; Qingjie Bi; Xiongwei Liu; Kuo Tian
err分享
err收藏
Continuous Probability Distribution Prediction of Image Emotions via Multitask Shared Sparse Regression
err2017-03-01
err156
PREAI
errZhao, Sicheng; Yao, Hongxun; Gao, Yue; Ji, Rongrong; Ding, Guiguang
err分享
err收藏
Adjuvant Systemic Chemotherapy vs Active Surveillance Following Up-front Resection of Isolated Synchronous Colorectal Peritoneal Metastases
err2020-08-13
err0
errOAAI
errKoen P. Rovers; Checca Bakkers; Felice N. van Erning; Jacobus W. A. Burger; Simon W. Nienhuijs; Geert A. A. M. Simkens; Geert-Jan M. Creemers; Patrick H. J. Hemmer; Cornelis J. A. Punt; Valery E. P. P. Lemmens; Pieter J. Tanis; Ignace H. J. T. de Hingh
err分享
err收藏
学者 查看更多内容