arrow
Return

dtoolAI: Reproducibility for Deep Learning

delete2020-08-01
delete23
delete
OA
AI
M
Matthew Hartley *
T
Tjelvar S. G. Olsson
DOI:10.1016/j.patter.2020.100073delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
Deep learning, a set of approaches using artificial neural networks, has generated rapid recent advancements in machine learning. Deep learning does, however, have the potential to reduce the reproducibility of scientific results. Model outputs are critically dependent on the data and processing approach used to initially generate the model, but this provenance information is usually lost during model training. To avoid a future reproducibility crisis, we need to improve our deep-learning model management. The FAIR principles for data stewardship and software/workflow implementation give excellent high-level guidance on ensuring effective reuse of data and software. We suggest some specific guidelines for the generation and use of deep-learning models in science and explain how these relate to the FAIR principles. We then present dtoolAI, a Python package that we have developed to implement these guidelines. The package implements automatic capture of provenance information during model training and simplifies model distribution.
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Patterns cover
Patterns
IF:
7.4
Papers:
935
Citations:
3.6K

Organization

U
uk research & innovation (ukri)
Scholars:
2.7W
Papers: 2.3W
Citations: 32