☆ 4.8 Article

NCBI reference sequences (RefSeq): a curated non-redundant sequence database of genomes, transcripts and proteins

NUCLEIC ACIDS RESEARCH (2007)

Journal

NUCLEIC ACIDS RESEARCH

Volume 35, Issue -, Pages D61-D65

Publisher

OXFORD UNIV PRESS

DOI: 10.1093/nar/gkl842

Keywords

Funding

NATIONAL LIBRARY OF MEDICINE [Z01LM000160] Funding Source: NIH RePORTER
Intramural NIH HHS Funding Source: Medline

Ask authors/readers for more resources

Protocol

Community support

Reagent

Community support

Abstract

NCBI's reference sequence (RefSeq) database (http://www.ncbi.nlm.nih.gov/RefSeq/) is a curated non-redundant collection of sequences representing genomes, transcripts and proteins. The database includes 3774 organisms spanning prokaryotes, eukaryotes and viruses, and has records for 2 879 860 proteins (RefSeq release 19). RefSeq records integrate information from multiple sources, when additional data are available from those sources and therefore represent a current description of the sequence and its features. Annotations include coding regions, conserved domains, tRNAs, sequence tagged sites (STS), variation, references, gene and protein product names, and database cross-references. Sequence is reviewed and features are added using a combined approach of collaboration and other input from the scientific community, prediction, propagation from GenBank and curation by NCBI staff. The format of all RefSeq records is validated, and an increasing number of tests are being applied to evaluate the quality of sequence and annotation, especially in the context of complete genomic sequence.

NCBI reference sequences (RefSeq): a curated non-redundant sequence database of genomes, transcripts and proteins

Journal

NUCLEIC ACIDS RESEARCH

Publisher

OXFORD UNIV PRESS

Keywords

Categories

Funding

Ask authors/readers for more resources

Protocol

Reagent

Authors

I am an author on this paper

Reviews

Primary Rating

Secondary Ratings

Novelty

Significance

Scientific rigor

Rate this paper

Recommended

NCBI reference sequences (RefSeq): a curated non-redundant sequence database of genomes, transcripts and proteins

Journal

NUCLEIC ACIDS RESEARCH

Publisher

OXFORD UNIV PRESS

Keywords

Categories

Funding

Ask authors/readers for more resources

Protocol

Reagent

Authors

I am an author on this paper

Reviews

Primary Rating

Secondary Ratings

Novelty

Significance

Scientific rigor

Rate this paper

Recommended

Export Citation

Share Paper