arrow
Return

Coefficient tree regression: fast, accurate and interpretable predictive modeling

delete2021-11-12
delete4
delete
OA
AI
Ö
Özge Sürer *
D
Daniel W. Apley
E
Edward C. Malthouse
DOI:10.1007/s10994-021-06091-7delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
The proliferation of data collection technologies often results in large data sets with many observations and many variables. In practice, highly relevant engineered features are often groups of predictors that share a common regression coefficient (i.e., the predictors in the group affect the response only via their collective sum), where the groups are unknown in advance and must be discovered from the data. We propose an algorithm called coefficient tree regression (CTR) to discover the group structure and fit the resulting regression model. In this regard CTR is an automated way of engineering new features, each of which is the collective sum of the predictors within each group. The algorithm can be used when the number of variables is larger than, or smaller than, the number of observations. Creating new features that affect the response in a similar manner improves predictive modeling, especially in domains where the relationships between predictors are not known a priori. CTR borrows computational strategies from both linear regression (fast model updating when adding/modifying a feature in the model) and regression trees (fast partitioning to form and split groups) to achieve outstanding computational and predictive performance. Finding features that represent hidden groups of predictors (i.e., a hidden ontology) that impact the response only via their sum also has major interpretability advantages, which we demonstrate with a real data example of predicting political affiliations with television viewing habits. In numerical comparisons over a variety of examples, we demonstrate that both computational expense and predictive performance are far superior to existing methods that create features as groups of predictors. Moreover, CTR has overall predictive performance that is comparable to or slightly better than the regular lasso method, which we include as a reference benchmark for comparison even though it is non-group-based, in addition to having substantial computational and interpretive advantages over lasso.
Keywords:
Aggregation
Group structure
Ontology
Feature engineering

Journal

Machine Learning cover
Machine Learning
IF:
2.9
Papers:
2.7K
Citations:
3.4W

Organization

N
Northwestern University
Scholars:
6.2W
Papers: 5.3W
Citations: 3.9K
Cited Papers

Cited Papers

IL-9 and Th9 cells: progress and challenges
err2013-09-11
err0
errOAAI
errPicheng Zhao; Xiang Xiao; Rafik M. Ghobrial; Xian C. Li
errShare
errSave
Parasuicide in Edinburgh—A Seven-Year Review 1968–74
err2018-01-29
err0
PREAI
errT. A. Holding; Dorothy Buglass; J. C. Duffy; Norman Kreitman
errShare
errSave
Random Matrices in Physics
err1967-01-01
err0
PREAI
errEugene P. Wigner
errShare
errSave
Isotopes of titanium in Aldebaran
err1977-01-01
err0
errOAAI
errD. L. Lambert; R. E. Luck
errShare
errSave
Expanding automotive electronic systems
err2002-01-01
err0
PREAI
errG. Leen; D. Heffernan
errShare
errSave
researcher View more