arrow
返回

Visual question generation for explicit questioning purposes based on target objects

delete2023-10-01
delete2
PRE
AI
J
Jiayuan Xie
J
Jiali Chen
W
Wenhao Fang
蔡毅 封面图
蔡毅 (Yi Cai) *
Prof. LI Qing 封面图
Prof. LI Qing (Qing Li)
DOI:10.1016/j.neunet.2023.08.007delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Visual question generation aims to focus on some target objects in an image to generate questions with certain questioning purposes. Existing studies mainly utilize an answer to extract the target object corresponding to the questioning purpose for questioning. However, answers fail to accurately and completely map to every target object, such as the objects corresponding to the answer are ambiguous or the answers are the relationship between multiple objects. To address this problem, we propose a content-controlled question generation model, which generates questions based on a given target object set specified from an image. Considering that the target objects have different contributions during the generation process, we design a recurrent generative architecture to explicitly control attention to different objects and their corresponding image information at each generative stage. Extensive experiments on the VQA v2.0 dataset and the Visual7w dataset show that the proposed model outperforms the state-of-the-art models and can controllably generate questions with specified content.(c) 2023 Elsevier Ltd. All rights reserved.
Keyword:
Visual question generation
Questioning purposes
Target object

期刊

Neural Networks 封面图
Neural Networks
IF:
6.3
论文数:
8.2K
被引数:
3.0W

机构

H
hong kong polytechnic university
学者数:
3.0W
论文数: 4.1W
被引数: 921
S
south china university of technology
学者数:
6.8W
论文数: 5.1W
被引数: 85
引用论文

引用论文

err
IF0
err
err0
PREAI
err
err分享
err收藏
err
IF0
err
err0
PREAI
err
err分享
err收藏
On polymorphism of 2-(4-fluorophenylamino)-5-(2,4-dihydroxybenzeno)-1,3,4-thiadiazole (FABT) DMSO solvates
err2013-01-01
err0
PREAI
errAnna A. Hoser; Daniel M. Kamiński; Arkadiusz Matwijczuk; Andrzej Niewiadomy; Mariusz Gagoś; Krzysztof Woźniak
err分享
err收藏
Dual Global Enhanced Transformer for image captioning
err2022-04-01
err63
PREAI
errXian, Tiantao; Li, Zhixin; Zhang, Canlong; Ma, Huifang
err分享
err收藏
Radial Graph Convolutional Network for Visual Question Generation
err2021-04-01
err39
PREAI
errXu, Xing; Wang, Tan; Yang, Yang; Hanjalic, Alan; Shen, Heng Tao
err分享
err收藏
学者 查看更多内容