☆ 4.7 Article

Amyloidogenic motifs revealed by n-gram analysis

SCIENTIFIC REPORTS (2017)

期刊

SCIENTIFIC REPORTS

卷 7, 期 -, 页码 -

出版社

NATURE PORTFOLIO

DOI: 10.1038/s41598-017-13210-9

关键词

类别

Multidisciplinary Sciences

资金

Wroclaw Center for Networking and Supercomputing [347]
KNOW Consortium
National Science Center [2015/17/N/NZ2/01845, 2017/24/T/NZ2/00003]

向作者/读者索取更多资源

Protocol

社区支持

Reagent

社区支持

摘要

Amyloids are proteins associated with several clinical disorders, including Alzheimer's, and Creutzfeldt-Jakob's. Despite their diversity, all amyloid proteins can undergo aggregation initiated by short segments called hot spots. To find the patterns defining the hot spots, we trained predictors of amyloidogenicity, using n-grams and random forest classifiers. Since the amyloidogenicity may not depend on the exact sequence of amino acids but on their more general properties, we tested 524,284 reduced amino acid alphabets of different lengths (three to six letters) to find the alphabet providing the best performance in cross-validation. The predictor based on this alphabet, called AmyloGram, was benchmarked against the most popular tools for the detection of amyloid peptides using an external data set and obtained the highest values of performance measures (AUC: 0.90, MCC: 0.63). Our results showed sequential patterns in the amyloids which are strongly correlated with hydrophobicity, a tendency to form beta-sheets, and lower flexibility of amino acid residues. Among the most informative n-grams of AmyloGram we identified 15 that were previously confirmed experimentally. AmyloGram is available as the web-server: http://smorfland.uni.wroc.pl/shiny/AmyloGram/ and as the R package AmyloGram. R scripts and data used to produce the results of this manuscript are available at http://github.com/michbur/AmyloGramAnalysis.

Amyloidogenic motifs revealed by n-gram analysis

期刊

SCIENTIFIC REPORTS

出版社

NATURE PORTFOLIO

关键词

类别

资金

向作者/读者索取更多资源

Protocol

Reagent

作者

我是这篇论文的作者

评论

主要评分

次要评分

新颖性

重要性

科学严谨性

评价这篇论文

推荐

Amyloidogenic motifs revealed by n-gram analysis

期刊

SCIENTIFIC REPORTS

出版社

NATURE PORTFOLIO

关键词

类别

资金

向作者/读者索取更多资源

Protocol

Reagent

作者

我是这篇论文的作者

评论

主要评分

次要评分

新颖性

重要性

科学严谨性

评价这篇论文

推荐

导出引文

分享论文