arrow
Return

Multi-modal learning with incomplete data

delete2026-08-31
delete0
delete
OA
AI
A
Alberto López *
J
John Zobolas
T
Tanguy Dumontier
T
Tero Aittokallio *
DOI:10.1038/s41467-026-77212-wdelete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Multi-modal learning, in which diverse data types are integrated and analyzed together, has become a central area of research in artificial intelligence, driving major advances in a wide range of domains. However, in many practical situations, certain modalities or variables may be missing for part of the samples, leading to a limited performance or failure of conventional methods. This has given a rise to the field of multi-modal learning with incomplete data, an area that has grown rapidly due to its broad real-world applications. Despite this, the community still lacks standardized tools to effectively handle incomplete multi-modal data. To fill this gap, we developed iMML, a unified, user-friendly Python package with versatile methods designed for integrating, processing, and analyzing incomplete multi-modal data. Successful use cases in biomedicine, text analysis, and computer vision for diverse machine learning tasks show the potency of iMML for making the best use of modern datasets in complex real-world applications. The iMML package is available at https://github.com/ocbe-uio/imml with an extensive documentation at https://imml.readthedocs.io/. Incomplete multi-modal datasets pose a major challenge for real-world machine learning applications. Here, authors present iMML, a unified open-source Python package designed to analyze and integrate incomplete multi-modal datasets for diverse machine learning tasks.

Journal

Nature Communications cover
Nature Communications
IF:
15.7
Papers:
9.3W
Citations:
91.2W

Organization

O
Oslo University Hospital
Scholars:
1.6K
Papers: 665
Citations: 1.9W
U
University of Oslo
Scholars:
3.6K
Papers: 1.6K
Citations: 0
Cited Papers

Cited Papers

No cited papers available