4.6 Article

COCOLASSO FOR HIGH-DIMENSIONAL ERROR-IN-VARIABLES REGRESSION

期刊

ANNALS OF STATISTICS
卷 45, 期 6, 页码 2400-2426

出版社

INST MATHEMATICAL STATISTICS
DOI: 10.1214/16-AOS1527

关键词

Convex optimization; error in variables; high-dimensional regression; LASSO; missing data

资金

  1. NSF [DMS-15-05111]

向作者/读者索取更多资源

Much theoretical and applied work has been devoted to high-dimensional regression with clean data. However, we often face corrupted data in many applications where missing data and measurement errors cannot be ignored. Loh and Wainwright [Ann. Statist. 40 (2012) 1637-1664] proposed a non-convex modification of the Lasso for doing high-dimensional regression with noisy and missing data. It is generally agreed that the virtues of convexity contribute fundamentally the success and popularity of the Lasso. In light of this, we propose a new method named CoCoLasso that is convex and can handle a general class of corrupted datasets. We establish the estimation error bounds of CoCoLasso and its asymptotic sign-consistent selection property. We further elucidate how the standard cross validation techniques can be misleading in presence of measurement error and develop a novel calibrated cross-validation technique by using the basic idea in CoCoLasso. The calibrated cross-validation has its own importance. We demonstrate the superior performance of our method over the nonconvex approach by simulation studies.

作者

我是这篇论文的作者
点击您的名字以认领此论文并将其添加到您的个人资料中。

评论

主要评分

4.6
评分不足

次要评分

新颖性
-
重要性
-
科学严谨性
-
评价这篇论文

推荐

暂无数据
暂无数据