arrow
Return

A Long-Tailed Image Classification Method Based on Enhanced Contrastive Visual Language

delete2023-07-26
delete0
delete
OA
AI
宋颖 cover
宋颖 (Ying Song) *
M
Mengxing Li
B
Bo Wang
DOI:10.3390/s23156694delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
To solve the problem that the common long-tailed classification method does not use the semantic features of the original label text of the image, and the difference between the classification accuracy of most classes and minority classes are large, the long-tailed image classification method based on enhanced contrast visual language trains the head class and tail class samples separately, uses text image to pre-train the information, and uses the enhanced momentum contrastive loss function and RandAugment enhancement to improve the learning of tail class samples. On the ImageNet-LT long-tailed dataset, the enhanced contrasting visual language-based long-tailed image classification method has improved all class accuracy, tail class accuracy, middle class accuracy, and the F-1 value by 3.4%, 7.6%, 3.5%, and 11.2%, respectively, compared to the BALLAD method. The difference in accuracy between the head class and tail class is reduced by 1.6% compared to the BALLAD method. The results of three comparative experiments indicate that the long-tailed image classification method based on enhanced contrastive visual language has improved the performance of tail classes and reduced the accuracy difference between the majority and minority classes.
Keywords:
long-tailed image classification
contrastive learning
data augmentation
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Sensors cover
Sensors
IF:
3.5
Papers:
7.1W
Citations:
20.9W

Organization

Z
Zhengzhou University of Light Industry
Scholars:
6.4K
Papers: 4.0K
Citations: 5.4K