arrow
Return

RefCap: image captioning with referent objects attributes

delete2023-12-07
delete3
delete
OA
AI
S
Seokmok Park
J
Joonki Paik *
DOI:10.1038/s41598-023-48916-6delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
In recent years, significant progress has been made in visual-linguistic multi-modality research, leading to advancements in visual comprehension and its applications in computer vision tasks. One fundamental task in visual-linguistic understanding is image captioning, which involves generating human-understandable textual descriptions given an input image. This paper introduces a referring expression image captioning model that incorporates the supervision of interesting objects. Our model utilizes user-specified object keywords as a prefix to generate specific captions that are relevant to the target object. The model consists of three modules including: (i) visual grounding, (ii) referring object selection, and (iii) image captioning modules. To evaluate its performance, we conducted experiments on the RefCOCO and COCO captioning datasets. The experimental results demonstrate that our proposed method effectively generates meaningful captions aligned with users' specific interests.
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Scientific Reports cover
Scientific Reports
IF:
3.9
Papers:
27.8W
Citations:
83.5W

Organization

C
Chung Ang University
Scholars:
1.3W
Papers: 1.4W
Citations: 133
Cited Papers

Cited Papers

Enhanced mechanical behaviour of lead zirconate titanate piezoelectric composites incorporating zinc oxide nanowhiskers
err2008-11-25
err0
PREAI
errLin Hai-Bo; Cao Mao-Sheng; Yuan Jie; Wang Da-Wei; Zhao Quan-Liang; Wang Fu-Chi
errShare
errSave
errShare
errSave
Visual Genome: Connecting Language and Vision Using Crowdsourced Dense Image Annotations
err2017-02-06
err3.1K
errOAAI
errKrishna, Ranjay; Zhu, Yuke; Groth, Oliver; Johnson, Justin; Hata, Kenji; Kravitz, Joshua; Chen, Stephanie; Kalantidis, Yannis; Li, Li-Jia; Shamma, David A.; Bernstein, Michael S.; Li Fei-Fei
errShare
errSave
Enhancing the alignment between target words and corresponding frames for video captioning
err2021-03-01
err41
PREAI
errTu, Yunbin; Zhou, Chang; Guo, Junjun; Gao, Shengxiang; Yu, Zhengtao
errShare
errSave
PTEN deletion is rare but often homogeneous in gastric cancer
err2012-05-25
err0
PREAI
errSormeh Mina; Benjamin A Bohn; Ronald Simon; Antje Krohn; Matthias Reeh; Dirk Arnold; Carsten Bokemeyer; Guido Sauter; Jakob R Izbicki; Andreas Marx; Phillip R Stahl
errShare
errSave
Necrotizing fasciitis caused by diabetic foot
err2021-02-01
err0
errOAAI
errZhengdong Zhang; Pan Liu; Banyin Yang; Jun Li; Wenzhao Wang; Hai Yang; Lei Liu
errShare
errSave
researcher View more