arrow
Return

Information-based clustering

delete2005-12-13
delete176
delete
OA
AI
N
Noam Slonim
G
Gurinder S. Atwal
G
Gašper Tkačik
W
William Bialek
DOI:10.1073/pnas.0507432102delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
In an age of increasingly large data sets, investigators in many different disciplines have turned to clustering as a tool for data analysis and exploration. Existing clustering methods, however, typically depend on several nontrivial assumptions about the structure of data. Here, we reformulate the clustering problem from an information theoretic perspective that avoids many of these assumptions. In particular, our formulation obviates the need for defining a cluster prototype, does not require an a priori. similarity metric, is invariant to changes in the representation of the data, and naturally captures nonlinear relations. We apply this approach to different domains and find that it consistently produces clusters that are more coherent than those extracted by existing algorithms. Finally, our approach provides a way of clustering based on collective notions of similarity rather than the traditional pairwise measures.
Keywords:
information theory
rate distortion
cluster analysis
gene expression
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

P
Proceedings of the National Academy of Sciences of the United States of America
IF:
9.1
Papers:
10.8W
Citations:
73.5W

Organization

No organization information available