arrow
返回

Quantifying relevance in learning and inference

delete2022-06-01
delete9
delete
OA
AI
M
Matteo Marsili
Y
Yasser Roudi *
DOI:10.1016/j.physrep.2022.03.001delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
Learning is a distinctive feature of intelligent behaviour. High-throughput experimental data and Big Data promise to open new windows on complex systems such as cells, the brain or our societies. Yet, the puzzling success of Artificial Intelligence and Machine Learning shows that we still have a poor conceptual understanding of learning. These applications push statistical inference into uncharted territories where data is high dimensional and scarce, and prior information on true models is scant if not totally absent. Here we review recent progress on understanding learning, based on the notion of relevance . The relevance, as we define it here, quantifies the amount of information that a dataset or the internal representation of a learning machine contains on the generative model of the data. This allows us to define maximally informative samples, on one hand, and optimal learning machines on the other. These are ideal limits of samples and of machines, that contain the maximal amount of information about the unknown generative process, at a given resolution (or level of compression). Both ideal limits exhibit critical features in the statistical sense: Maximally informative samples are characterised by a power-law frequency distribution (statistical criticality) and optimal learning machines by an anomalously large susceptibility. The trade-off between resolution (i.e. compression) and relevance distinguishes the regime of noisy representations from that of lossy compression. These are separated by a special point characterised by Zipf's law statistics. This identifies samples obeying Zipf's law as the most compressed loss-less representations that are optimal in the sense of maximal relevance. Criticality in optimal learning machines manifests in an exponential degeneracy of energy levels, that leads to unusual thermodynamic properties. This distinctive feature is consistent with the invariance of the classification under coarse graining of the output, which is a desirable property of learning machines. This theoretical framework is corroborated by empirical analysis showing (i) how the concept of relevance can be useful to identify relevant variables in high-dimensional inference and (ii) that widely used machine learning architectures approach reasonably well the ideal limit of optimal learning machines, within the limits of the data with which they are trained. (c) 2022 The Authors. Published by Elsevier B.V.
Keyword:
Relevance
Statistical inference
Machine learning
Information theory
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

P
Physics Reports-Review Section of Physics Letters
IF:
29.5
论文数:
3.0K
被引数:
3.8W

机构

暂无机构信息
引用论文

引用论文

Engaging tacit knowledge in support of organizational learning
errVINE
IF0
err2008-04-11
err0
PREAI
errDavid Bennet; Alex Bennet
err分享
err收藏
err分享
err收藏
Amino acid derived sulfonamide hydroxamates as inhibitors of procollagen C-Proteinase: solid-Phase synthesis of ornithine analogues
err2001-08-01
err0
PREAI
errSharon M. Dankwardt; Robert L. Martin; Christine S. Chan; Harold E. Van Wart; Keith A.M. Walker; Nancy G. Delaet; Leslie A. Robinson
err分享
err收藏
C4d as a significant predictor for humoral rejection in renal allografts
err2005-08-24
err0
PREAI
errChen Jianghua; Xie Wenqing; Wang Huiping; Jin Juan; Wu Jianyong; He Qiang
err分享
err收藏
Internal Models in Physics Problem Solving
err2009-12-14
err0
PREAI
errYuichiro Anzai; Tohru Yokoyama
err分享
err收藏
Polycystic kidneys and del (4)(q21.1q21.3): further delineation of a distinct phenotype
err2005-01-01
err0
PREAI
errM. Velinov; J. Kupferman; H. Gu; M.J. Macera; A. Babu; E.C. Jenkins; G. Kupchik
err分享
err收藏
学者 查看更多内容