arrow
Return

Multimodal Deep Learning Framework for Skin Lesion Classification

delete2026-01-01
delete0
PRE
AI
B
Balaji Banothu *
G
Ghassan Hamarneh
S
S. Nickolas
G
Gaurav Patil
DOI:10.1007/978-3-031-93697-5_15delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Medical imaging plays a crucial role in modern healthcare, serving as a fundamental component for diagnosis and treatment. The selection of imaging modalities for specific diagnostic tasks often involves balancing accessibility, cost, and performance. In this study, we propose a novel approach that utilizes knowledge from a standard modality characterized by higher performance but lower feasibility to enhance the effectiveness of a meta-modality, which is more accessible yet underperforming. Our focus is on applying deep learning techniques in medical image diagnosis by building a lightweight model. This lightweight mapping model utilizes potential representations leveraging the standard modality to improve the training of a model that relies solely on meta-modality. The effectiveness of our approach are illustrated through its use in a clinical setting for multi-task classification of skin lesions, relying on both clinical and dermoscopic images. Our results indicate a notable enhancement within the diagnostic accuracy of the metamodality, attained without the need for the standard modality during inference. The balanced accuracy increased from 0.87 for clinical images alone to 0.92 for the combined model using both modalities. Furthermore, the experiments were performed repeatedly to ensure their consistency under several random weight configurations.
Keywords:
Multimodal learning
Knowledge Distillation Network
Student-Teacher learning
Classification
Dermascopic Images

Journal

C
COMPUTER VISION AND IMAGE PROCESSING, CVIP 2024, PT IV
IF:
0
Papers:
32
Citations:
0

Organization

N
national institute of technology (nit system)
Scholars:
4.0W
Papers: 3.7W
Citations: 31