返回
Boosted multivariate trees for longitudinal data
DOI:10.1007/s10994-016-5597-1.png)
摘要
En 中文
Machine learning methods provide a powerful approach for analyzing longitudinal data in which repeated measurements are observed for a subject over time. We boost multivariate trees to fit a novel flexible semi-nonparametric marginal model for longitudinal data. In this model, features are assumed to be nonparametric, while feature-time interactions are modeled semi-nonparametrically utilizing P-splines with estimated smoothing parameter. In order to avoid overfitting, we describe a relatively simple in sample cross-validation method which can be used to estimate the optimal boosting iteration and which has the surprising added benefit of stabilizing certain parameter estimates. Our new multivariate tree boosting method is shown to be highly flexible, robust to covariance misspecification and unbalanced designs, and resistant to overfitting in high dimensions. Feature selection can be used to identify important features and feature-time interactions. An application to longitudinal data of forced 1-second lung expiratory volume (FEV1) for lung transplant patients identifies an important feature-time interaction and illustrates the ease with which our method can find complex relationships in longitudinal data.
Keyword:
Gradient boosting
Marginal model
Multivariate regression tree
P-splines
Smoothing parameter
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
2.9
论文数:
2.7K
被引数:
3.4W
机构
引用论文
Highly efficient electrocatalytic hydrogen production by nickel promoted molybdenum sulfide microspheres catalysts镍促进硫化钼微球催化剂的高效电催化制氢
RSC Advances
IF0
Preparation of graphite-loaded epoxy-based voltammetric electrodes using a multi-layer coating and hardening technique
The Analyst
IF0

