返回
Unsupervised Visual Anomaly Detection Using Self-Supervised Pre-Trained Transformer
DOI:10.1109/ACCESS.2024.3454753.png)
摘要
En 中文
In the various industrial manufacturing processes, the automatic visual inspection system is an essential part as it reduces the chances of delivering defective products and the cost of training and hiring experts for manual inspection. In this work, we propose a new unsupervised anomaly detection method inspired by the masked language model for the automatic visual inspection system. The proposed method consists of an image tokenizer and two subnetworks, a reconstruction subnetwork, and a segmentation subnetwork. We adopt a pre-trained self-supervised vision Transformer model to use it as an image tokenizer. Our first subnetwork is trained to predict the anomaly-free patch tokens and the second subnetwork is trained to produce anomaly segmentation results from both the reconstructed and input patch tokens. During training, only the two subnetworks are optimized, and parameters of an image tokenizer are frozen. Experimental results show that the proposed method exhibits better performance than conventional methods in detecting defective products by achieving 99.05% I-AUROC on MVTecAD dataset and 94.8% I-AUROC on BTAD.
Keyword:
Image reconstruction
Image segmentation
Transformers
Computational modeling
Location awareness
Feature extraction
Anomaly detection
Data augmentation
Self-supervised learning
data-augmentation
self-supervised learning
transformer
期刊
IF:
3.6
论文数:
9.8W
被引数:
29.4W
机构
引用论文
Inflexibility of mental planning: A characteristic disorder with prefrontal lobe lesions?心理计划的僵化: 前额叶病变的特征性障碍?
First-Principles Computation of the Vibrational Entropy of Ordered and DisorderedNi3Al有序和无序振动熵的第一性原理计算 Ni3Al
Deep Architecture for High-Speed Railway Insulator Surface Defect Detection: Denoising Autoencoder With Multitask Learning高速铁路绝缘子表面缺陷检测的深度架构: 多任务学习去噪自编码器

