arrow
返回

Deep learning-based clustering approaches for bioinformatics

delete2020-02-01
delete156
delete
OA
AI
M
Md. Rezaul Karim *
O
Oya Beyan
A
Achille Zappa
I
Ivan G. Costa
D
Dietrich Rebholz‐Schuhmann
M
Michael Cochez
S
Stefan Decker
DOI:10.1093/bib/bbz170delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
Clustering is central to many data-driven bioinformatics research and serves a powerful computational method. In particular, clustering helps at analyzing unstructured and high-dimensional data in the form of sequences, expressions, texts and images. Further, clustering is used to gain insights into biological processes in the genomics level, e.g. clustering of gene expressions provides insights on the natural structure inherent in the data, understanding gene functions, cellular processes, subtypes of cells and understanding gene regulations. Subsequently, clustering approaches, including hierarchical, centroid-based, distribution-based, density-based and self-organizing maps, have long been studied and used in classical machine learning settings. In contrast, deep learning (DL)-based representation and feature learning for clustering have not been reviewed and employed extensively. Since the quality of clustering is not only dependent on the distribution of data points but also on the learned representation, deep neural networks can be effective means to transform mappings from a high-dimensional data space into a lower-dimensional feature space, leading to improved clustering results. In this paper, we review state-of-the-art DL-based approaches for cluster analysis that are based on representation learning, which we hope to be useful, particularly for bioinformatics research. Further, we explore in detail the training procedures of DL-based clustering algorithms, point out different clustering quality metrics and evaluate several DL-based approaches on three bioinformatics use cases, including bioimaging, cancer genomics and biomedical text mining. We believe this review and the evaluation results will provide valuable insights and serve a starting point for researchers wanting to apply DL-based unsupervised methods to solve emerging bioinformatics research problems.
Keyword:
GENE
CANCER
REPRESENTATION
MODEL
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Briefings in Bioinformatics 封面图
Briefings in Bioinformatics
IF:
7.7
论文数:
5.8K
被引数:
2.7W

机构

R
RWTH Aachen University
学者数:
3.5W
论文数: 2.6W
被引数: 3.6W
F
fraunhofer institute center schloss birlinghoven
学者数:
196
论文数: 151
被引数: 0
U
university of zurich
学者数:
5.1W
论文数: 4.0W
被引数: 65
O
ollscoil na gaillimhe-university of galway
学者数:
1.1W
论文数: 8.7K
被引数: 5
F
fraunhofer gesellschaft
学者数:
1.6W
论文数: 1.2W
被引数: 24
学者 查看更多机构
引用论文

引用论文

err分享
err收藏
Multi-omics approaches for precision obesity management
err2023-01-30
err0
errOAAI
errSelam Woldemariam; Thomas E. Dorner; Thomas Wiesinger; Katharina Viktoria Stein
err分享
err收藏
River Flow 2004
err
IF0
err2004-06-15
err0
PREAI
err
err分享
err收藏
学者 查看更多内容