NCBI reference sequences (RefSeq): a curated non-redundant sequence database of genomes, transcripts and proteins
National Institutes of Health · National Center for Biotechnology Information
Abstract
NCBI's reference sequence (RefSeq) database (http://www.ncbi.nlm.nih.gov/RefSeq/) is a curated non-redundant collection of sequences representing genomes, transcripts and proteins. The database includes 3774 organisms spanning prokaryotes, eukaryotes and viruses, and has records for 2,879,860 proteins (RefSeq release 19). RefSeq records integrate information from multiple sources, when additional data are available from those sources and therefore represent a current description of the sequence and its features. Annotations include coding regions, conserved domains, tRNAs, sequence tagged sites (STS), variation, references, gene and protein product names, and database cross-references. Sequence is reviewed and…
Citation impact
- FWCI
- 115.59
- Percentile
- 100%
- References
- 22
Authors
3Topics & keywords
- RefSeq
- GenBank
- Biology
- Sequence database
- Sequence (biology)
- Genome
- Annotation
- Ensembl