arrow
返回

On explaining machine learning models by evolving crucial and compact features

delete2020-03-01
delete19
delete
OA
AI
M
Marco Virgolin *
T
Tanja Alderliesten
P
Peter A. N. Bosman
DOI:10.1016/j.swevo.2019.100640delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
Feature construction can substantially improve the accuracy of Machine Learning (ML) algorithms. Genetic Programming (GP) has been proven to be effective at this task by evolving non-linear combinations of input features. GP additionally has the potential to improve ML explainability since explicit expressions are evolved. Yet, in most GP works the complexity of evolved features is not explicitly bound or minimized though this is arguably key for explainability. In this article, we assess to what extent GP still performs favorably at feature construction when constructing features that are (1) Of small-enough number, to enable visualization of the behavior of the ML model; (2) Of small-enough size, to enable interpretability of the features themselves; (3) Of sufficient informative power, to retain or even improve the performance of the ML algorithm. We consider a simple feature construction scheme using three different GP algorithms, as well as random search, to evolve features for five ML algorithms, including support vector machines and random forest. Our results on 21 datasets pertaining to classification and regression problems show that constructing only two compact features can be sufficient to rival the use of the entire original feature set. We further find that a modern GP algorithm, GP-GOMEA, performs best overall. These results, combined with examples that we provide of readable constructed features and of 2D visualizations of ML behavior, lead us to positively conclude that GP-based feature construction still works well when explicitly searching for compact features, making it extremely helpful to explain ML models.
Keyword:
Feature construction
Interpretable machine learning
Genetic programming
GOMEA
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Swarm and Evolutionary Computation 封面图
Swarm and Evolutionary Computation
IF:
8.5
论文数:
2.2K
被引数:
1.0W

机构

U
university of amsterdam
学者数:
6.0W
论文数: 5.1W
被引数: 94
D
Delft University of Technology
学者数:
2.6W
论文数: 2.5W
被引数: 3.8W
引用论文

引用论文

err分享
err收藏
err分享
err收藏
Multi-objective genetic programming for feature extraction and data visualization
err2015-10-26
err32
errOAAI
errCano, Alberto; Ventura, Sebastian; Cios, Krzysztof J.
err分享
err收藏
err分享
err收藏
学者 查看更多内容