☆ 4.6 Article

APPLES: Scalable Distance-Based Phylogenetic Placement with or without Alignments

SYSTEMATIC BIOLOGY (2020)

期刊

SYSTEMATIC BIOLOGY

卷 69, 期 3, 页码 566-578

出版社

OXFORD UNIV PRESS

DOI: 10.1093/sysbio/syz063

关键词

Distance-based methods; genome skimming; phylogenetic placement

类别

Evolutionary Biology

资金

National Science Foundation (NSF) [IIS-1565862]
National Institutes of Health (NIH) [5P30AI027767-28]
NSF [ACI-1053575, NSF-1815485]

向作者/读者索取更多资源

Protocol

社区支持

Reagent

社区支持

摘要

Placing a new species on an existing phylogeny has increasing relevance to several applications. Placement can be used to update phylogenies in a scalable fashion and can help identify unknown query samples using (meta-)barcoding, skimming, or metagenomic data. Maximum likelihood (ML) methods of phylogenetic placement exist, but these methods are not scalable to reference trees with many thousands of leaves, limiting their ability to enjoy benefits of dense taxon sampling in modern reference libraries. They also rely on assembled sequences for the reference set and aligned sequences for the query. Thus, ML methods cannot analyze data sets where the reference consists of unassembled reads, a scenario relevant to emerging applications of genome skimming for sample identification. We introduce APPLES, a distance-based method for phylogenetic placement. Compared to ML, APPLES is an order of magnitude faster and more memory efficient, and unlike ML, it is able to place on large backbone trees (tested for up to 200,000 leaves). We show that using dense references improves accuracy substantially so that APPLES on dense trees is more accurate than ML on sparser trees, where it can run. Finally, APPLES can accurately identify samples without assembled reference or aligned queries using kmer-based distances, a scenario that ML cannot handle. APPLES is available publically at github.com/balabanmetin/apples.

APPLES: Scalable Distance-Based Phylogenetic Placement with or without Alignments

期刊

SYSTEMATIC BIOLOGY

出版社

OXFORD UNIV PRESS

关键词

类别

资金

向作者/读者索取更多资源

Protocol

Reagent

作者

我是这篇论文的作者

评论

主要评分

次要评分

新颖性

重要性

科学严谨性

评价这篇论文

推荐

APPLES: Scalable Distance-Based Phylogenetic Placement with or without Alignments

期刊

SYSTEMATIC BIOLOGY

出版社

OXFORD UNIV PRESS

关键词

类别

资金

向作者/读者索取更多资源

Protocol

Reagent

作者

我是这篇论文的作者

评论

主要评分

次要评分

新颖性

重要性

科学严谨性

评价这篇论文

推荐

导出引文

分享论文