arrow
Return

Automatic image captioning system using a deep learning approach

delete2023-05-27
delete1
PRE
AI
G
Gerard Deepak *
S
Sowmya Gali
A
Abhilash Sonker
B
Bobin Cherian Jos
K
K. V. Daya Sagar
C
Charanjeet Singh
DOI:10.1007/s00500-023-08544-8delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
This paper's residual network is tailored to increase the high-quality image caption generation ability. The captioning is exploited using the relevant content with high-quality interpretation. The research develops a Residual Attention Generative Adversarial Network (RAGAN) and uses attention-based residual learning in Generative Adversarial Network (GAN) to improve the diversity and fidelity of the generated image captions. The RAGAN exploits the words based on the feature maps faster to generate high-quality captions. The RAGAN improves the diversity of captions generated and increases the language metrics scores. The generator is designed as an encoder-decoder mechanism that operates in an unsupervised manner. The residual learning is adopted between the encoder and decoder network. The discriminator is connected to a language evaluator unit, which provides feed-forward to the generator and discriminator to either positively or negatively influence the image captioning process. The experiments show that the proposed RAGAN performs better than the state-of-the-art GAN models.
Keywords:
Image captioning
Deep learning
Generative adversarial network
Residual learning

Journal

Soft Computing cover
Soft Computing
IF:
2.5
Papers:
1.0W
Citations:
2.1W

Organization

M
mar athanasius college of engineering
Scholars:
35
Papers: 21
Citations: 0