GenBank

Nucleic Acids Res. 2014 Jan;42(Database issue):D32-7. doi: 10.1093/nar/gkt1030. Epub 2013 Nov 11.

Abstract

GenBank is a comprehensive database that contains publicly available nucleotide sequences for over 280,000 formally described species. These sequences are obtained primarily through submissions from individual laboratories and batch submissions from large-scale sequencing projects, including whole-genome shotgun and environmental sampling projects. Most submissions are made using the web-based BankIt or standalone Sequin programs, and GenBank staff assign accession numbers upon data receipt. Daily data exchange with the European Nucleotide Archive and the DNA Data Bank of Japan ensures worldwide coverage. GenBank is accessible through the National Center for Biotechnology Information (NCBI) Entrez retrieval system, which integrates data from the major DNA and protein sequence databases along with taxonomy, genome, mapping, protein structure and domain information, and the biomedical journal literature via PubMed. BLAST provides sequence similarity searches of GenBank and other sequence databases. Complete bimonthly releases and daily updates of the GenBank database are available by FTP. To access GenBank and its related retrieval and analysis services, begin at the NCBI home page: www.ncbi.nlm.nih.gov.

Publication types

  • Research Support, N.I.H., Intramural

MeSH terms

  • Bacteria / classification
  • Bacteria / genetics
  • Databases, Nucleic Acid*
  • High-Throughput Nucleotide Sequencing
  • Internet
  • Molecular Sequence Annotation
  • Sequence Analysis, DNA*