arrow
返回

Deep Learning for Image-to-Text Generation A technical overview

delete2017-11-01
delete72
PRE
AI
X
Xiaodong He *
L
Li Deng
DOI:10.1109/MSP.2017.2741510delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Generating a natural language description from an image is an emerging interdisciplinary problem at the intersection of computer vision, natural language processing, and artificial intelligence ( AI). This task, often referred to as image or visual captioning, forms the technical foundation of many important applications, such as semantic visual search, visual intelligence in chatting robots, photo and video sharing in social media, and aid for visually impaired people to perceive surrounding visual content. Thanks to the recent advances in deep learning, the AI research community has witnessed tremendous progress in visual captioning in recent years. In this article, we will first summarize this exciting emerging visual captioning area. We will then analyze the key development and the major progress the community has made, their impact in both research and industry deployment, and what lies ahead in future breakthroughs.
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

IEEE Signal Processing Magazine 封面图
IEEE Signal Processing Magazine
IF:
9.6
论文数:
1.1W
被引数:
1.7W

机构

M
Microsoft
学者数:
3.0K
论文数: 2.7K
被引数: 7
引用论文

引用论文

Nerve growth factor and its receptors TrkA and p75 are upregulated in the brain of mdx dystrophic mouse
err2009-07-01
err0
PREAI
errB. Nico; D. Mangieri; A. De Luca; P. Corsi; V. Benagiano; R. Tamma; T. Annese; V. Longo; E. Crivellato; D. Ribatti
err分享
err收藏
Growth Factors in the Intestinal Tract
err2018-01-01
err0
PREAI
errMichael A. Schumacher; Soula Danopoulos; Denise Al Alam; Mark R. Frey
err分享
err收藏
err分享
err收藏
学者 查看更多内容