arrow
Return

Incorporating object counts into remote sensing image captioning

delete2024-08-22
delete11
delete
OA
AI
Z
Zihao Ni
宗兆云 (Zhaoyun Zong)
P
Peng Ren *
DOI:10.1080/17538947.2024.2392847delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Existing methods for remote sensing image captioning tend to describe a remote sensing image using generic language that lacks specific information about object counts. To address this limitation, we propose a novel framework for generating a caption that includes object count information for the remote sensing image. Our proposed framework comprises three modules: object counting, preliminary captioning, and numeral editing. The object counting module identifies objects in a remote sensing image and determines object counts. The preliminary captioning module generates a caption that may lack object count information. The numeral editing module incorporates the object counts into the caption, resulting in a more precise caption. Our proposed framework outperforms existing methods, as demonstrated through evaluations on three remote sensing image datasets. Our proposed framework is a significant step toward more precise and informative remote sensing image captioning.
Keywords:
Remote sensing
earth observation
artificial intelligence
image processing

Journal

International Journal of Digital Earth cover
International Journal of Digital Earth
IF:
4.9
Papers:
1.9K
Citations:
4.7K

Organization

C
china university of petroleum
Scholars:
4.1W
Papers: 2.7W
Citations: 30