返回
Local descriptor-based multi-prototype network for few-shot Learning
DOI:10.1016/j.patcog.2021.107935.png)
摘要
En 中文
Prototype-based few-shot learning methods are promising in that they are simple yet effective to handle any-shot problems, and many prototype associated works are raised since then. However, these traditional prototype-based methods generally use only one single prototype to represent a class, which essentially cannot effectively estimate the complicated distribution of a class. To tackle this problem, we propose a novel Local descriptor-based Multi-Prototype Network (LMPNet) in this paper, a well-designed framework that generates an embedding space with multiple prototypes. Specifically, the proposed LMPNet employs local descriptors to represent each image, which can capture more informative and subtler cues of an image than the normally adopted image-level features. Moreover, to alleviate the uncertainty introduced by the fixed construction (averaging over samples) of prototypes, we introduce a channel squeeze and spatial excitation (sSE) attention module to learn multiple local descriptor-based prototypes for each class through end-to-end learning. Extensive experiments on both few-shot and fine-grained few-shot image classification tasks have been conducted on various benchmark datasets, including miniImageNet, tieredImageNet, Stanford Dogs, Stanford Cars, and CUB-200-2010. The experimental results of our LMPNet on above datasets show tangibly learning performance improvements and distinguishable outcomes over the baseline models. (c) 2021 Elsevier Ltd. All rights reserved.
Keyword:
Few-shot learning
Image classification
Local descriptors
Multiple prototypes
End-to-end learning
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
7.6
论文数:
1.3W
被引数:
4.5W
机构
引用论文
Scheduled sampling for one-shot learning via matching network通过匹配网络进行一次性学习的计划采样
PATTERN RECOGNITION
IF7.6
Grad-CAM: Visual Explanations from Deep Networks via Gradient-Based LocalizationGrad-cam: 通过基于梯度的定位从深度网络进行视觉解释

