arrow
返回

Extending Contrastive Learning to Unsupervised Coreset Selection

delete2022-01-01
delete9
delete
OA
AI
J
Jeongwoo Ju
H
Heechul Jung
Y
Yoonju Oh
J
Junmo Kim *
DOI:10.1109/ACCESS.2022.3142758delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Self-supervised contrastive learning offers a means of learning informative features from a pool of unlabeled data. In this paper, we investigate another useful approach. We propose an entirely unlabeled coreset selection method. In this regard, contrastive learning, one of several self-supervised methods, was recently proposed and has consistently delivered the highest performance. This prompted us to choose two leading methods for contrastive learning: the simple framework for contrastive learning of visual representations (SimCLR) and the momentum contrastive (MoCo) learning framework. We calculated the cosine similarities for each example of an epoch for the entire duration of the contrastive learning process and subsequently accumulated the cosine similarity values to obtain the coreset score. Our assumption was that a sample with low similarity would likely behave as a coreset. Compared with existing coreset selection methods with labels, our approach reduced the cost associated with human annotation. In this study, the unsupervised method implemented for coreset selection achieved improvements of 1.25% (for CIFAR10), 0.82% (for SVHN), and 0.19% (for QMNIST) over a randomly selected subset with a size of 30%. Furthermore, our results are comparable to those of the existing supervised coreset selection methods. The differences between the proposed and the above mentioned supervised coreset selection method (forgetting events) were 0.81% on the CIFAR10 dataset, -2.08% on the SVHN dataset (the proposed method outperformed the existing method), and 0.01% on the QMNIST dataset at a subset size of 30%. In addition, our proposed approach exhibited robustness even if the coreset selection model and target model were not identical (e.g., using ResNet18 as a selection model and ResNet101 as the target model). Lastly, we obtained more concrete proof that our coreset examples are highly informative by showing the performance gap between the coreset and non-coreset samples in the coreset cross test experiment. We observed a pair of performance ((testing: non-coreset, training: coreset), (testing: coreset, training: non-coreset)), i.e. (94.27%, 67.39 %) for CIFAR10, (98.24%, 83.30%) for SVHN, and (99.89%, 93.07%) for QMNIST with a subset size of 30%.
Keyword:
Task analysis
Training
Measurement
Annotations
Licenses
Feature extraction
Deep learning
Coreset selection
image classification
self-supervised learning
contrastive learning

期刊

IEEE Access 封面图
IEEE Access
IF:
3.6
论文数:
9.8W
被引数:
29.4W

机构

K
kyungpook national university (knu)
学者数:
1.8W
论文数: 1.8W
被引数: 14
引用论文

引用论文

err分享
err收藏
Feature Selective Projection with Low-Rank Embedding and Dual Laplacian Regularization
err2019-01-01
err141
PREAI
errTang, Chang; Liu, Xinwang; Zhu, Xinzhong; Xiong, Jian; Li, Miaomiao; Xia, Jingyuan; Wang, Xiangke; Wang, Lizhe
err分享
err收藏
Synthesis of n-type semiconducting diamond film using diphosphorus pentaoxide as the doping source以五氧化二磷为掺杂源合成n型半导体金刚石膜
err1990-10-01
err0
PREAI
errKen Okano; Hideo Kiyota; Tatsuya Iwasaki; Yoshitaka Nakamura; Yukio Akiba; Tateki Kurosu; Masamori Iida; Terutaro Nakamura
err分享
err收藏
LOW: Training deep neural networks by learning optimal sample weights
err2021-02-01
err20
PREAI
errSantiago, Carlos; Barata, Catarina; Sasdelli, Michele; Carneiro, Gustavo; Nascimento, Jacinto C.
err分享
err收藏
err分享
err收藏
Automated Pupil Perimetry Pupil Field Mapping in Patients and Normal Subjects
err1991-04-01
err0
PREAI
errRandy H. Kardon; Pinar Aydin Kirkali; H. Stanley Thompson
err分享
err收藏
err分享
err收藏
Learning Cooperative Personalized Policies from Gaze Data
err2019-10-17
err0
PREAI
errChristoph Gebhardt; Brian Hecox; Bas van Opheusden; Daniel Wigdor; James Hillis; Otmar Hilliges; Hrvoje Benko
err分享
err收藏
学者 查看更多内容