4.6 Article

Sparse partial least squares regression for simultaneous dimension reduction and variable selection

出版社

WILEY
DOI: 10.1111/j.1467-9868.2009.00723.x

关键词

Chromatin immuno-precipitation; Dimension reduction; Gene expression; Lasso; Microarrays; Partial least squares; Sparsity; Variable and feature selection

资金

  1. National Institutes of Health [H6003747]
  2. National Science Foundation [DMS 0804597]

向作者/读者索取更多资源

Partial least squares regression has been an alternative to ordinary least squares for handling multicollinearity in several areas of scientific research since the 1960s. It has recently gained much attention in the analysis of high dimensional genomic data. We show that known asymptotic consistency of the partial least squares estimator for a univariate response does not hold with the very large p and small n paradigm. We derive a similar result for a multivariate response regression with partial least squares. We then propose a sparse partial least squares formulation which aims simultaneously to achieve good predictive performance and variable selection by producing sparse linear combinations of the original predictors. We provide an efficient implementation of sparse partial least squares regression and compare it with well-known variable selection and dimension reduction approaches via simulation experiments. We illustrate the practical utility of sparse partial least squares regression in a joint analysis of gene expression and genomewide binding data.

作者

我是这篇论文的作者
点击您的名字以认领此论文并将其添加到您的个人资料中。

评论

主要评分

4.6
评分不足

次要评分

新颖性
-
重要性
-
科学严谨性
-
评价这篇论文

推荐

暂无数据
暂无数据