arrow
Return

Instance spaces for machine learning classification

delete2017-12-28
delete84
delete
OA
AI
M
Mario Andrés Muñoz
L
Laura Villanova
D
Davaatseren Baatar
K
Kate Smith‐Miles *
DOI:10.1007/s10994-017-5629-5delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
This paper tackles the issue of objective performance evaluation of machine learning classifiers, and the impact of the choice of test instances. Given that statistical properties or features of a dataset affect the difficulty of an instance for particular classification algorithms, we examine the diversity and quality of the UCI repository of test instances used by most machine learning researchers. We show how an instance space can be visualized, with each classification dataset represented as a point in the space. The instance space is constructed to reveal pockets of hard and easy instances, and enables the strengths and weaknesses of individual classifiers to be identified. Finally, we propose a methodology to generate new test instances with the aim of enriching the diversity of the instance space, enabling potentially greater insights than can be afforded by the current UCI repository.
Keywords:
Classification
Meta-learning
Test data
Instance space
Performance evaluation
Algorithm footprints
Test instance generation
Instance difficulty
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Machine Learning cover
Machine Learning
IF:
2.9
Papers:
2.6K
Citations:
3.4W

Organization

M
Monash University
Scholars:
5.4W
Papers: 5.4W
Citations: 79