☆ 4.8 Article

Use of simulated data sets to evaluate the fidelity of metagenomic processing methods

NATURE METHODS (2007)

期刊

NATURE METHODS

卷 4, 期 6, 页码 495-500

出版社

NATURE PUBLISHING GROUP

DOI: 10.1038/nmeth1043

关键词

类别

Biochemical Research Methods

向作者/读者索取更多资源

Protocol

社区支持

Reagent

社区支持

摘要

Metagenomics is a rapidly emerging field of research for studying microbial communities. To evaluate methods presently used to process metagenomic sequences, we constructed three simulated data sets of varying complexity by combining sequencing reads randomly selected from 113 isolate genomes. These data sets were designed to model real metagenomes in terms of complexity and phylogenetic composition. We assembled sampled reads using three commonly used genome assemblers (Phrap, Arachne and JAZZ), and predicted genes using two popular gene-finding pipelines (fgenesb and CRITICA/GLIMMER). The phylogenetic origins of the assembled contigs were predicted using one sequence similarity-based ( blast hit distribution) and two sequence composition-based (PhyloPythia, oligonucleotide frequencies) binning methods. We explored the effects of the simulated community structure and method combinations on the fidelity of each processing step by comparison to the corresponding isolate genomes. The simulated data sets are available online to facilitate standardized benchmarking of tools for metagenomic analysis.

Use of simulated data sets to evaluate the fidelity of metagenomic processing methods

期刊

NATURE METHODS

出版社

NATURE PUBLISHING GROUP

关键词

类别

向作者/读者索取更多资源

Protocol

Reagent

作者

我是这篇论文的作者

评论

主要评分

次要评分

新颖性

重要性

科学严谨性

评价这篇论文

推荐

Use of simulated data sets to evaluate the fidelity of metagenomic processing methods

期刊

NATURE METHODS

出版社

NATURE PUBLISHING GROUP

关键词

类别

向作者/读者索取更多资源

Protocol

Reagent

作者

我是这篇论文的作者

评论

主要评分

次要评分

新颖性

重要性

科学严谨性

评价这篇论文

推荐

导出引文

分享论文