☆ 4.4 Article

A fractal method to distinguish coding and non-coding sequences in a complete genome based on a number sequence representation

JOURNAL OF THEORETICAL BIOLOGY (2005)

期刊

JOURNAL OF THEORETICAL BIOLOGY

卷 232, 期 4, 页码 559-567

出版社

ACADEMIC PRESS LTD- ELSEVIER SCIENCE LTD

DOI: 10.1016/j.jtbi.2004.09.002

关键词

coding sequences; non-coding sequences; DNA

类别

Biology Mathematical & Computational Biology

向作者/读者索取更多资源

Protocol

社区支持

Reagent

社区支持

摘要

A fractal method to distinguish coding and non-coding sequences in a complete genome is proposed, based on different statistical behaviors between these two kinds of sequences. We first propose a number sequence representation of DNA sequences. Multifractal analysis is then performed on the measure representation of the obtained number sequence. The three exponents C-1, C-1 and C-2 are selected from the result of multifractal analysis. Each DNA may be represented by a point in the three-dimensional space generated by these three-component vectors. It is shown that points corresponding to coding and non-coding sequences in the complete genome of many prokaryotes are roughly distributed in different regions. Fisher's discriminant algorithm can be used to separate these two regions in the spanned space. If the point (C-1,C-1,C-2) for a DNA sequence is situated in the region corresponding to coding sequences, the sequence is discriminated as a coding sequence; otherwise, the sequence is classified as a non-coding one. For all 51 prokaryotes we considered, the average discriminant accuracies p(c), p(nc), q(c), and q(nc), reach 72.28%, 84.65%, 72.53% and 84.18%, respectively. (C) 2004 Elsevier Ltd. All rights reserved.

A fractal method to distinguish coding and non-coding sequences in a complete genome based on a number sequence representation

期刊

JOURNAL OF THEORETICAL BIOLOGY

出版社

ACADEMIC PRESS LTD- ELSEVIER SCIENCE LTD

关键词

类别

向作者/读者索取更多资源

Protocol

Reagent

作者

我是这篇论文的作者

评论

主要评分

次要评分

新颖性

重要性

科学严谨性

评价这篇论文

推荐

A fractal method to distinguish coding and non-coding sequences in a complete genome based on a number sequence representation

期刊

JOURNAL OF THEORETICAL BIOLOGY

出版社

ACADEMIC PRESS LTD- ELSEVIER SCIENCE LTD

关键词

类别

向作者/读者索取更多资源

Protocol

Reagent

作者

我是这篇论文的作者

评论

主要评分

次要评分

新颖性

重要性

科学严谨性

评价这篇论文

推荐

导出引文

分享论文