arrow
Return

The randomized information coefficient: assessing dependencies in noisy data

delete2017-09-19
delete9
delete
OA
AI
S
Simone Romano *
N
Nguyễn Xuân Vinh
K
Karin Verspoor
J
James Bailey
DOI:10.1007/s10994-017-5664-2delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
When differentiating between strong and weak relationships using information theoretic measures, the variance plays an important role: the higher the variance, the lower the chance to correctly rank the relationships. We propose the randomized information coefficient (RIC), a mutual information based measure with low variance, to quantify the dependency between two sets of numerical variables. We first formally establish the importance of achieving low variance when comparing relationships using the mutual information estimated with grids. Second, we experimentally demonstrate the effectiveness of RIC for (i) detecting noisy dependencies and (ii) ranking dependencies for the applications of genetic network inference and feature selection for regression. Across these tasks, RIC is very competitive over other 16 state-of-the-art measures. Other prominent features of RIC include its simplicity and efficiency, making it a promising new method for dependency assessment.
Keywords:
Dependency measures
Noisy relationships
Normalized mutual information
Randomized ensembles
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Machine Learning cover
Machine Learning
IF:
2.9
Papers:
2.6K
Citations:
3.4W

Organization

U
university of melbourne
Scholars:
5.7W
Papers: 5.4W
Citations: 69