arrow
返回

Image captioning with transformer and knowledge graph

delete2021-03-01
delete63
PRE
AI
张
张宇 (Yu Zhang) *
X
Xinyu Shi
米思娅 封面图
米思娅 (Siya Mi)
X
Xu Yang
DOI:10.1016/j.patrec.2020.12.020delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
The Transformer model has achieved very good results in machine translation tasks. In this paper, we adopt the Transformer model for the image captioning task. To promote the performance of image captioning, we improve the Transformer model from two aspects. First, we augment the maximum likelihood estimation (MLE) with an extra Kullback-Leibler (KL) divergence term to distinguish the difference between incorrect predictions. Second, we introduce a method to help the Transformer model generate captions by leveraging the knowledge graph. Experiments on benchmark datasets demonstrate the effectiveness of our method. (c) 2021 Elsevier B.V. All rights reserved.
Keyword:
Image captioning
Transformer
Knowledge graph
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Pattern Recognition Letters 封面图
Pattern Recognition Letters
IF:
3.3
论文数:
7.9K
被引数:
1.6W

机构

N
Nanyang Technological University
学者数:
4.9W
论文数: 4.8W
被引数: 8.1W
S
southeast university - china
学者数:
5.3W
论文数: 4.9W
被引数: 57
引用论文

引用论文

“It's almost like they're trying to hide it”: How User-Provided Image Descriptions Have Failed to Make Twitter Accessible
err2019-05-13
err0
errOAAI
errCole Gleason; Patrick Carrington; Cameron Cassidy; Meredith Ringel Morris; Kris M. Kitani; Jeffrey P. Bigham
err分享
err收藏
err分享
err收藏
A Comprehensive Survey of Deep Learning for Image Captioning用于图像字幕的深度学习综述
err2019-02-04
err421
errOAAI
errHossain, Md Zakir; Sohel, Ferdous; Shiratuddin, Mohd Fairuz; Laga, Hamid
err分享
err收藏
Image caption generation with high-level image features
err2019-05-01
err44
PREAI
errDing, Songtao; Qu, Shiru; Xi, Yuling; Sangaiah, Arun Kumar; Wan, Shaohua
err分享
err收藏
Image Caption Generation with Part of Speech Guidance
err2019-03-01
err55
PREAI
errHe, Xinwei; Shi, Baoguang; Bai, Xiang; Xia, Gui-Song; Zhang, Zhaoxiang; Dong, Weisheng
err分享
err收藏
Leveraging unpaired out -of -domain data for image captioning利用未配对的域外数据进行图像字幕
err2020-04-01
err15
PREAI
errChen, Xinghan; Zhang, Mingxing; Wang, Zheng; Zuo, Lin; Li, Bo; Yang, Yang
err分享
err收藏
没有更多内容