4.4 Article

GenFamClust: an accurate, synteny-aware and reliable homology inference algorithm

期刊

BMC EVOLUTIONARY BIOLOGY
卷 16, 期 -, 页码 -

出版社

BMC
DOI: 10.1186/s12862-016-0684-2

关键词

Homology inference; Gene synteny; Gene similarity; Gene family; Clustering; Gene order conservation

资金

  1. Higher Education Commission (HEC) of Pakistan
  2. EuroSPIN (Erasmus Mundus joint doctoral program)
  3. Swedish e-Science Research Center (SeRC)

向作者/读者索取更多资源

Background: Homology inference is pivotal to evolutionary biology and is primarily based on significant sequence similarity, which, in general, is a good indicator of homology. Algorithms have also been designed to utilize conservation in gene order as an indication of homologous regions. We have developed GenFamClust, a method based on quantification of both gene order conservation and sequence similarity. Results: In this study, we validate GenFamClust by comparing it to well known homology inference algorithms on a synthetic dataset. We applied several popular clustering algorithms on homologs inferred by GenFamClust and other algorithms on a metazoan dataset and studied the outcomes. Accuracy, similarity, dependence, and other characteristics were investigated for gene families yielded by the clustering algorithms. GenFamClust was also applied to genes from a set of complete fungal genomes and gene families were inferred using clustering. The resulting gene families were compared with a manually curated gold standard of pillars from the Yeast Gene Order Browser. We found that the gene-order component of GenFamClust is simple, yet biologically realistic, and captures local synteny information for homologs. Conclusions: The study shows that GenFamClust is a more accurate, informed, and comprehensive pipeline to infer homologs and gene families than other commonly used homology and gene-family inference methods.

作者

我是这篇论文的作者
点击您的名字以认领此论文并将其添加到您的个人资料中。

评论

主要评分

4.4
评分不足

次要评分

新颖性
-
重要性
-
科学严谨性
-
评价这篇论文

推荐

暂无数据
暂无数据