arrow
返回

Cross-Modal Data Programming Enables Rapid Medical Machine Learning

delete2020-05-01
delete31
delete
OA
AI
J
Jared Dunnmon *
A
Alexander Ratner
K
Khaled Saab
N
Nishith Khandwala
M
Matthew Markert
H
Hersh Sagreiya
R
Roger E. Goldman
C
Christopher Lee‐Messer
M
Matthew P. Lungren
D
Daniel L. Rubin
C
Christopher Ré
DOI:10.1016/j.patter.2020.100019delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
A major bottleneck in developing clinically impactful machine learning models is a lack of labeled training data for model supervision. Thus, medical researchers increasingly turn to weaker, noisier sources of supervision, such as leveraging extractions from unstructured text reports to supervise image classification. A key challenge in weak supervision is combining sources of information that may differ in quality and have correlated errors. Recently, a statistical theory of weak supervision called data programming has shown promise in addressing this challenge. Data programming now underpinsmany deployed machine-learning systems in the technology industry, even for critical applications. We propose a new technique for applying data programming to the problem of cross-modal weak supervision in medicine, wherein weak labels derived from an auxiliary modality (e.g., text) are used to train models over a different target modality (e.g., images). We evaluate our approach on diverse clinical tasks via direct comparison to institution-scale, hand-labeled data-sets. We find that our supervision technique increases model performance by up to 6 points area under the receiver operating characteristic curve (ROC-AUC) over baseline methods by improving both coverage and quality of the weak labels. Our approach yields models that on average perform within 1.75 points ROC-AUC of those supervised with physician-years of hand labeling and outperform those supervised with physician-months of hand labeling by 10.25 points ROC-AUC, while using only person-days of developer time and clinician work-a time saving of 96%. Our results suggest that modern weak supervision techniques such as data programming may enable more rapid development and deployment of clinically useful machine-learning models.
Keyword:
DEEP
ALGORITHM
CANCER
TIMES
EEG
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Patterns 封面图
Patterns
IF:
7.4
论文数:
949
被引数:
3.6K

机构

S
Stanford University
学者数:
9.6W
论文数: 8.2W
被引数: 17.0W
引用论文

引用论文

The cellular mode of action of the anti-epileptic drug 5, 5-diphenylhydantoin
err1978-03-01
err0
PREAI
errJ. GAVIN PERRY; LESLIE MCKINNEY; PAUL DE WEER
err分享
err收藏
err分享
err收藏
Time is brain - Quantified
errSTROKE
IF8.9
err2006-01-01
err1.5K
errOAAI
errSaver, JL
err分享
err收藏
Deep convolutional neural network for the automated detection and diagnosis of seizure using EEG signals
err2018-09-01
err1.1K
PREAI
errAcharya, U. Rajendra; Oh, Shu Lih; Hagiwara, Yuki; Tan, Jen Hong; Adeli, Hojjat
err分享
err收藏
Global hydro-environmental lake characteristics at high spatial resolution
err2022-06-23
err0
errOAAI
errBernhard Lehner; Mathis L. Messager; Maartje C. Korver; Simon Linke
err分享
err收藏
Predicting non-small cell lung cancer prognosis by fully automated microscopic pathology image features全自动显微病理图像特征预测非小细胞肺癌预后
err2016-08-16
err710
errOAAI
errYu, Kun-Hsing; Zhang, Ce; Berry, Gerald J.; Altman, Russ B.; Re, Christopher; Rubin, Daniel L.; Snyder, Michael
err分享
err收藏
Deep learning for chest radiograph diagnosis: A retrospective comparison of the CheXNeXt algorithm to practicing radiologists胸部x光片诊断的深度学习: CheXNeXt算法与实践放射科医生的回顾性比较
err2018-11-20
err728
errOAAI
errRajpurkar, Pranav; Irvin, Jeremy; Ball, Robyn L.; Zhu, Kaylie; Yang, Brandon; Mehta, Hershel; Duan, Tony; Ding, Daisy; Bagul, Aarti; Langlotz, Curtis P.; Patel, Bhavik N.; Yeom, Kristen W.; Shpanskaya, Katie; Blankenberg, Francis G.; Seekins, Jayne; Amrhein, Timothy J.; Mong, David A.; Halabi, Safwan S.; Zucker, Evan J.; Ng, Andrew Y.; Lungren, Matthew P.
err分享
err收藏
An explainable deep-learning algorithm for the detection of acute intracranial haemorrhage from small datasets
err2018-12-17
err304
PREAI
errLee, Hyunkwang; Yune, Sehyo; Mansouri, Mohammad; Kim, Myeongchan; Tajmir, Shahein H.; Guerrieri, Claude E.; Ebert, Sarah A.; Pomerantz, Stuart R.; Romero, Javier M.; Kamalian, Shahmir; Gonzalez, Ramon G.; Lev, Michael H.; Do, Synho
err分享
err收藏
学者 查看更多内容