☆ 4.7 Article

A fast data-driven method for genotype imputation, phasing and local ancestry inference: Mendellmpute.jl

BIOINFORMATICS (2021)

期刊

BIOINFORMATICS

卷 37, 期 24, 页码 4756-4763

出版社

OXFORD UNIV PRESS

DOI: 10.1093/bioinformatics/btab489

关键词

类别

Biochemical Research Methods Biotechnology & Applied Microbiology Computer Science, Interdisciplinary Applications Mathematical & Computational Biology Statistics & Probability

资金

NIH [R01-HG009120, T32-HG002536, R01-HG006139, R01-GM053275, R35GM141798]
NSF [DMS-2054253]

向作者/读者索取更多资源

Protocol

社区支持

Reagent

社区支持

智能总结 New
摘要

A novel data-mining method for genotype imputation and phasing was introduced, utilizing efficient linear algebra routines for calculations and delivering similar prediction accuracy with better memory usage and faster run-times compared to existing methods. The method operates on both dosage data and unphased genotype data, imputing missing genotypes and phasing untyped SNPs simultaneously.

Motivation: Current methods for genotype imputation and phasing exploit the volume of data in haplotype reference panels and rely on hidden Markov models (HMMs). Existing programs all have essentially the same imputation accuracy, are computationally intensive and generally require prephasing the typed markers. Results: We introduce a novel data-mining method for genotype imputation and phasing that substitutes highly efficient linear algebra routines for HMM calculations. This strategy, embodied in our Julia program Mendellmpute.jl, avoids explicit assumptions about recombination and population structure while delivering similar prediction accuracy, better memory usage and an order of magnitude or better run-times compared to the fastest competing method. Mendellmpute operates on both dosage data and unphased genotype data and simultaneously imputes missing genotypes and phase at both the typed and untyped SNPs (single nucleotide polymorphisms). Finally, Mendellmpute naturally extends to global and local ancestry estimation and lends itself to new strategies for data compression and hence faster data transport and sharing.

A fast data-driven method for genotype imputation, phasing and local ancestry inference: Mendellmpute.jl

期刊

BIOINFORMATICS

出版社

OXFORD UNIV PRESS

关键词

类别

资金

向作者/读者索取更多资源

Protocol

Reagent

作者

我是这篇论文的作者

评论

主要评分

次要评分

新颖性

重要性

科学严谨性

评价这篇论文

推荐

A fast data-driven method for genotype imputation, phasing and local ancestry inference: Mendellmpute.jl

期刊

BIOINFORMATICS

出版社

OXFORD UNIV PRESS

关键词

类别

资金

向作者/读者索取更多资源

Protocol

Reagent

作者

我是这篇论文的作者

评论

主要评分

次要评分

新颖性

重要性

科学严谨性

评价这篇论文

推荐

导出引文

分享论文