arrow
返回

A novel automatic image caption generation using bidirectional long-short term memory framework

delete2021-04-19
delete10
PRE
AI
Z
Zhongfu Ye *
R
Rashid Khan
N
Nuzhat Naqvi
DOI:10.1007/s11042-021-10632-6delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Image Captioning, the process of generating a textual description of an image, has emerged as a hot research due to its practical importance in many domains. It is a challenging task as it uses both Natural Language Processing and Computer Vision related fields to generate the captions. Despite the fact that the literature has reported notable image captioning methodologies, they still lag in accomplishing the substantial performance level for diverse datasets. This paper proposes an image caption generating mechanism based on Optimized Bidirectional Long Short-Term Memory (B-LSTM) model. We propose a variant of Moth Flame Optimization (PMFO), termed here as Proposed Moth Flame Optimization (PMFO), which has logarithmic spiral update based on correlation. The performance of the proposed model is demonstrated on benchmark datasets like Flicker 8 k, Flicker30k, VizWik and COCO datasets using renowned metrics such as CIDEr, BLEU, SPICE and ROUGH. The performance analysis proves that the B-LSTM achieves better performance on caption generation than state-of-the-art methods.
Keyword:
Image captioning
inception v3
B-LSTM
P-MFO optimization
Bleu score
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Multimedia Tools and Applications 封面图
Multimedia Tools and Applications
IF:
3
论文数:
2.0W
被引数:
3.2W

机构

C
chinese academy of sciences
学者数:
56.7W
论文数: 45.0W
被引数: 704
引用论文

引用论文

Stearoyl-CoA Desaturase-1 (SCD1) Augments Saturated Fatty Acid-Induced Lipid Accumulation and Inhibits Apoptosis in Cardiac Myocytes
err2012-03-08
err0
errOAAI
errHiroki Matsui; Tomoyuki Yokoyama; Kenichi Sekiguchi; Daisuke Iijima; Hiroaki Sunaga; Moeno Maniwa; Manabu Ueno; Tatsuya Iso; Masashi Arai; Masahiko Kurabayashi
err分享
err收藏
3G structure for image caption generation
err2019-02-01
err29
errOAAI
errYuan, Aihong; Li, Xuelong; Lu, Xiaoqiang
err分享
err收藏
err分享
err收藏
Modeling visual and word-conditional semantic attention for image captioning
err2018-09-01
err14
PREAI
errWu, Chunlei; Wei, Yiwei; Chu, Xiaoliang; Su, Fei; Wang, Leiquan
err分享
err收藏
学者 查看更多内容