Cross-validated bagged learning

详细信息查看全文

作者：Maya L. Petersen ; Annette M. Molinaro ; S ; ra E. Sinisi ; Mark J. van der Laan
关键词：Bootstrap aggregation ; Data-adaptive regression ; Resistant HIV ; Deletion/Substitution/Addition algorithm
刊名：Journal of Multivariate Analysis
出版年：2007
出版时间：October 2007
年：2007
卷：98
期：9
页码：1693-1704
全文大小：185 K

文摘

Many applications aim to learn a high dimensional parameter of a data generating distribution based on a sample of independent and identically distributed observations. For example, the goal might be to estimate the conditional mean of an outcome given a list of input variables. In this prediction context, bootstrap aggregating (bagging) has been introduced as a method to reduce the variance of a given estimator at little cost to bias. Bagging involves applying an estimator to multiple bootstrap samples and averaging the result across bootstrap samples. In order to address the curse of dimensionality, a common practice has been to apply bagging to estimators which themselves use cross-validation, thereby using cross-validation within a bootstrap sample to select fine-tuning parameters trading off bias and variance of the bootstrap sample-specific candidate estimators. In this article we point out that in order to achieve the correct bias variance trade-off for the parameter of interest, one should apply the cross-validation selector externally to candidate bagged estimators indexed by these fine-tuning parameters. We use three simulations to compare the new cross-validated bagging method with bagging of cross-validated estimators and bagging of non-cross-validated estimators.

地址：北京市海淀区学院路29号邮编：100083

电话：办公室：(+86 10)66554848；文献借阅、咨询服务、科技查新：66554700