返回
Interpretable machine learning for genomics
DOI:10.1007/s00439-021-02387-9.png)
摘要
En 中文
High-throughput technologies such as next-generation sequencing allow biologists to observe cell function with unprecedented resolution, but the resulting datasets are too large and complicated for humans to understand without the aid of advanced statistical methods. Machine learning (ML) algorithms, which are designed to automatically find patterns in data, are well suited to this task. Yet these models are often so complex as to be opaque, leaving researchers with few clues about underlying mechanisms. Interpretable machine learning (iML) is a burgeoning subdiscipline of computational statistics devoted to making the predictions of ML models more intelligible to end users. This article is a gentle and critical introduction to iML, with an emphasis on genomic applications. I define relevant concepts, motivate leading methodologies, and provide a simple typology of existing approaches. I survey recent examples of iML in genomics, demonstrating how such techniques are increasingly integrated into research workflows. I argue that iML solutions are required to realize the promise of precision medicine. However, several open challenges remain. I examine the limitations of current state-of-the-art tools and propose a number of directions for future research. While the horizon for iML in genomics is wide and bright, continued progress requires close collaboration across disciplines.
Keyword:
FALSE DISCOVERY RATE
CAUSAL INFERENCE
BLACK-BOX
VARIABLE-SELECTION
BREAST-CANCER
EXPLANATIONS
DECISIONS
HEALTH
AI
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
3.6
论文数:
4.6K
被引数:
8.9K
机构
引用论文
Degradation of reactive dyes I. A comparative study of ozonation, enzymatic and photochemical processes活性染料的降解I.臭氧、酶和光化学过程的比较研究
Chemosphere
IF0
Dissecting racial bias in an algorithm used to manage the health of populations在用于管理人群健康的算法中剖析种族偏见
SCIENCE
IF45.8

