4.5 Article

Structural zeros in high-dimensional data with applications to microbiome studies

期刊

BIOSTATISTICS
卷 18, 期 3, 页码 422-433

出版社

OXFORD UNIV PRESS
DOI: 10.1093/biostatistics/kxw053

关键词

Classification; High dimension; Microbiome data; Missing data; Sparsity

资金

  1. Intramural Research Program of the NIH, NIEHS [Z01 ES101744-04]
  2. Israeli Science Foundation [1256/13]

向作者/读者索取更多资源

This paper is motivated by the recent interest in the analysis of high-dimensional microbiome data. A key feature of these data is the presence of structural zeros which are microbes missing from an observation vector due to an underlying biological process and not due to error in measurement. Typical notions of missingness are unable to model these structural zeros. We define a general framework which allows for structural zeros in the model and propose methods of estimating sparse high-dimensional covariance and precision matrices under this setup. We establish error bounds in the spectral and Frobenius norms for the proposed estimators and empirically verify them with a simulation study. The proposed methodology is illustrated by applying it to the global gut microbiome data ofYatsunenko and others (2012. Human gut microbiome viewed across age and geography. Nature 486, 222-227). Using our methodology we classify subjects according to the geographical location on the basis of their gut microbiome.

作者

我是这篇论文的作者
点击您的名字以认领此论文并将其添加到您的个人资料中。

评论

主要评分

4.5
评分不足

次要评分

新颖性
-
重要性
-
科学严谨性
-
评价这篇论文

推荐

暂无数据
暂无数据