arrow
返回

Model selection and error estimation without the agonizing pain

delete2018-03-22
delete12
PRE
AI
L
Luca Oneto *
DOI:10.1002/widm.1252delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
How can we select the best performing data-driven model? How can we rigorously estimate its generalization error? Statistical learning theory (SLT) answers these questions by deriving nonasymptotic bounds on the generalization error of a model or, in other words, by delivering upper bounding of the true error of the learned model based just on quantities computed on the available data. However, for a long time, SLT has been considered only as an abstract theoretical framework, useful for inspiring new learning approaches, but with limited applicability to practical problems. The purpose of this review is to give an intelligible overview of the problems of model selection (MS) and error estimation (EE), by focusing on the ideas behind the different SLT-based approaches and simplifying most of the technical aspects with the purpose of making them more accessible and usable in practice. We start by presenting the seminal works of the 80s until the most recent results, then discuss open problems and finally outline future directions of this field of research. This article is categorized under: Technologies > Statistical Fundamentals Algorithmic Development > Statistics
Keyword:
algorithmic stability
bootstrap
compression bound
cross validation
differential privacy
error estimation
in-sample methods
model selection
out-of-sample methods
PAC-Bayes
Rademacher complexity theory
union bound
Vapnik-Chervonenkis theory
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Wiley Interdisciplinary Reviews-Data Mining and Knowledge Discovery 封面图
Wiley Interdisciplinary Reviews-Data Mining and Knowledge Discovery
IF:
11.7
论文数:
546
被引数:
5.3K

机构

U
university of genoa
学者数:
3.0W
论文数: 2.2W
被引数: 20
引用论文

引用论文

err分享
err收藏
err分享
err收藏
学者 查看更多内容