4.7 Article

Interaction-based feature selection and classification for high-dimensional biological data

期刊

BIOINFORMATICS
卷 28, 期 21, 页码 2834-2842

出版社

OXFORD UNIV PRESS
DOI: 10.1093/bioinformatics/bts531

关键词

-

资金

  1. Hong Kong Research Grant Council [642207, 601312]
  2. NIH [R01 GM070789, GM070789-0551]
  3. NSF [DMS-0714669]

向作者/读者索取更多资源

Motivation: Epistasis or gene-gene interaction has gained increasing attention in studies of complex diseases. Its presence as an ubiquitous component of genetic architecture of common human diseases has been contemplated. However, the detection of gene-gene interaction is difficult due to combinatorial explosion. Results: We present a novel feature selection method incorporating variable interaction. Three gene expression datasets are analyzed to illustrate our method, although it can also be applied to other types of high-dimensional data. The quality of variables selected is evaluated in two ways: first by classification error rates, then by functional relevance assessed using biological knowledge. We show that the classification error rates can be significantly reduced by considering interactions. Secondly, a sizable portion of genes identified by our method for breast cancer metastasis overlaps with those reported in gene-to-system breast cancer (G2SBC) database as disease associated and some of them have interesting biological implication. In summary, interaction-based methods may lead to substantial gain in biological insights as well as more accurate prediction.

作者

我是这篇论文的作者
点击您的名字以认领此论文并将其添加到您的个人资料中。

评论

主要评分

4.7
评分不足

次要评分

新颖性
-
重要性
-
科学严谨性
-
评价这篇论文

推荐

暂无数据
暂无数据