arrow
返回

QUANTILE REGRESSION UNDER MEMORY CONSTRAINT

delete2019-12-01
delete126
delete
OA
AI
X
Xi Chen *
刘卫东 (Weidong Liu)
Y
Yichen Zhang
DOI:10.1214/18-AOS1777delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
This paper studies the inference problem in quantile regression (QR) for a large sample size n but under a limited memory constraint, where the memory can only store a small batch of data of size m. A natural method is the naive divide-and-conquer approach, which splits data into batches of size m, computes the local QR estimator for each batch and then aggregates the estimators via averaging. However, this method only works when n = o(m(2)) and is computationally expensive. This paper proposes a computationally efficient method, which only requires an initial QR estimator on a small batch of data and then successively refines the estimator via multiple rounds of aggregations. Theoretically, as long as n grows polynomially in m, we establish the asymptotic normality for the obtained estimator and show that our estimator with only a few rounds of aggregations achieves the same efficiency as the QR estimator computed on all the data. Moreover, our result allows the case that the dimensionality p goes to infinity. The proposed method can also be applied to address the QR problem under distributed computing environment (e.g., in a large-scale sensor network) or for real-time streaming data.
Keyword:
Quantile regression
sample quantile
divide-and-conquer
distributed inference
streaming data

期刊

Annals of Statistics 封面图
Annals of Statistics
IF:
3.7
论文数:
2.8K
被引数:
2.9W

机构

S
shanghai jiao tong university
学者数:
15.7W
论文数: 11.7W
被引数: 159
N
New York University
学者数:
4.4W
论文数: 3.9W
被引数: 5.8W