GenBank

Nucleic Acids Res. 2011 Jan;39(Database issue):D32-7. doi: 10.1093/nar/gkq1079. Epub 2010 Nov 10.

Abstract

GenBank® is a comprehensive database that contains publicly available nucleotide sequences for more than 380,000 organisms named at the genus level or lower, obtained primarily through submissions from individual laboratories and batch submissions from large-scale sequencing projects, including whole genome shotgun (WGS) and environmental sampling projects. Most submissions are made using the web-based BankIt or standalone Sequin programs, and accession numbers are assigned by GenBank staff upon receipt. Daily data exchange with the European Nucleotide Archive (ENA) and the DNA Data Bank of Japan (DDBJ) ensures worldwide coverage. GenBank is accessible through the NCBI Entrez retrieval system that integrates data from the major DNA and protein sequence databases along with taxonomy, genome, mapping, protein structure and domain information, and the biomedical journal literature via PubMed. BLAST provides sequence similarity searches of GenBank and other sequence databases. Complete bimonthly releases and daily updates of the GenBank database are available by FTP. To access GenBank and its related retrieval and analysis services, begin at the NCBI Homepage: www.ncbi.nlm.nih.gov.

Publication types

  • Research Support, N.I.H., Intramural

MeSH terms

  • Databases, Nucleic Acid*
  • Expressed Sequence Tags
  • Genomics
  • High-Throughput Nucleotide Sequencing
  • Metagenomics
  • Molecular Sequence Annotation
  • Software