arrow
返回

Composite Object Relation Modeling for Few-Shot Scene Recognition

delete2023-01-01
delete1
PRE
AI
宋
宋新航 (Xinhang Song)
C
Chenlong Liu
H
Haitao Zeng
Y
Yaohui Zhu
G
Gongwei Chen
X
Xiaorong Qin
S
Shuqiang Jiang *
DOI:10.1109/TIP.2023.3321475delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
The goal of few-shot image recognition is to classify different categories with only one or a few training samples. Previous works of few-shot learning mainly focus on simple images, such as object or character images. Those works usually use a convolutional neural network (CNN) to learn the global image representations from training tasks, which are then adapted to novel tasks. However, there are many more abstract and complex images in real world, such as scene images, consisting of many object entities with flexible spatial relations among them. In such cases, global features can hardly obtain satisfactory generalization ability due to the large diversity of object relations in the scenes, which may hinder the adaptability to novel scenes. This paper proposes a composite object relation modeling method for few-shot scene recognition, capturing the spatial structural characteristic of scene images to enhance adaptability on novel scenes, considering that objects commonly co- occurred in different scenes. In different few-shot scene recognition tasks, the objects in the same images usually play different roles. Thus we propose a task-aware region selection module (TRSM) to further select the detected regions in different few-shot tasks. In addition to detecting object regions, we mainly focus on exploiting the relations between objects, which are more consistent to the scenes and can be used to cleave apart different scenes. Objects and relations are used to construct a graph in each image, which is then modeled with graph convolutional neural network. The graph modeling is jointly optimized with few-shot recognition, where the loss of few-shot learning is also capable of adjusting graph based representations. Typically, the proposed graph based representations can be plugged in different types of few-shot architectures, such as metric-based and meta-learning methods. Experimental results of few-shot scene recognition show the effectiveness of the proposed method.
Keyword:
Scene recognition
few-shot learning
graph modeling
generalization ability

期刊

IEEE Transactions on Image Processing 封面图
IEEE Transactions on Image Processing
IF:
13.7
论文数:
1.0W
被引数:
8.4W

机构

C
chinese academy of sciences
学者数:
56.7W
论文数: 45.0W
被引数: 704
引用论文

引用论文

Ingestion of corrosive acids
err1989-09-01
err0
PREAI
errShowkat Ali Zargar; Rakesh Kochhar; Birender Nagi; Saroj Mehta; Satish Kumar Mehta
err分享
err收藏
SUN Database: Exploring a Large Collection of Scene CategoriesSUN数据库: 探索大量场景类别
err2014-08-13
err181
PREAI
errXiao, Jianxiong; Ehinger, Krista A.; Hays, James; Torralba, Antonio; Oliva, Aude
err分享
err收藏
Sight-threatening Keratopathy Complicating Anti-TNF Therapy in Crohnʼs Disease
err2014-01-01
err0
errOAAI
errFederica Fasci-Spurio; Alexandra Thompson; Stephen Madill; Peter Koay; David Mansfield; Jack Satsangi
err分享
err收藏
学者 查看更多内容