NCBI (GenBank / RefSeq / dbSNP)
The National Center for Biotechnology Information (NCBI) is one of the world’s most widely used resources for accessing biological and genomic information. It provides comprehensive DNA and RNA sequence data through GenBank, curated reference sequences through RefSeq, and genetic variation information through databases such as dbSNP.
NCBI integrates genomic data with scientific literature, gene and protein information, genome browsers, and a wide range of bioinformatics tools. Researchers use these resources for sequence analysis, genome annotation, comparative genomics, variant identification, evolutionary studies, and literature-based research.
Key Resources:
- GenBank β Public repository of nucleotide sequences.
- RefSeq β Curated reference sequences for genomes, genes, transcripts, and proteins.
- dbSNP β Database of single nucleotide polymorphisms (SNPs) and other genetic variants.
Website: https://www.ncbi.nlm.nih.gov/
EMBL-EBI / ENA (European Nucleotide Archive)
The European Nucleotide Archive (ENA), maintained by the European Molecular Biology Laboratory β European Bioinformatics Institute (EMBL-EBI), is one of the world’s primary repositories for nucleotide sequence data. It stores a wide variety of genomic information, including raw sequencing reads, assembled genomes, transcript sequences, and annotated nucleotide records.
ENA supports the submission, storage, retrieval, and analysis of sequencing data generated by researchers worldwide. It is widely used in genomics, transcriptomics, metagenomics, evolutionary biology, and population genetics research.
ENA is a founding member of the International Nucleotide Sequence Database Collaboration (INSDC), alongside GenBank (NCBI) and DDBJ, ensuring synchronized sharing of nucleotide sequence data among the three major international databases.
Website: https://www.ebi.ac.uk/ena
DDBJ (DNA Data Bank of Japan)
The DNA Data Bank of Japan (DDBJ) is Japan’s primary public repository for nucleotide sequence information. It collects, archives, and provides access to DNA and RNA sequence data submitted by researchers from around the world.
As a member of the International Nucleotide Sequence Database Collaboration (INSDC), DDBJ exchanges sequence data daily with GenBank and ENA, ensuring that submitted nucleotide sequences are accessible through all three major international databases.
Researchers use DDBJ to search nucleotide sequences, access genome-related datasets, and submit newly generated sequencing data. The database supports research in molecular biology, genomics, evolutionary biology, biodiversity, and large-scale sequencing projects.
Website: https://www.ddbj.nig.ac.jp/
CNGBdb (China National GeneBank DataBase)
CNGBdb (China National GeneBank DataBase) is a comprehensive biological data platform developed by the China National GeneBank (CNGB) to support the storage, management, sharing, and analysis of large-scale genomic and multi-omics datasets.
The database provides access to diverse biological data, including genomic, transcriptomic, proteomic, metabolomic, microbiome, and other multi-omics datasets. In addition to data storage, CNGBdb offers computational tools for data visualization, exploration, annotation, and bioinformatics analysis.
CNGBdb is designed to facilitate large-scale life science research and supports studies in genomics, precision medicine, agriculture, biodiversity, population genetics, and bioinformatics by promoting open scientific data sharing and collaboration.
Website: https://db.cngb.org/
The Arabidopsis Information Resource (TAIR)
The Arabidopsis Information Resource (TAIR) is the primary database for the model plant Arabidopsis thaliana. It provides comprehensive genomic, genetic, molecular, and functional information, including genome sequences, gene annotations, gene expression data, metabolic pathways, and mutant resources.
TAIR is widely used in plant genetics, molecular biology, functional genomics, crop improvement, and plant biotechnology research. It serves as a reference resource for understanding gene function and plant biological processes.
Researchers use TAIR to:
- Access genome sequences and gene annotations
- Study gene function and regulation
- Explore gene expression and mutant information
- Analyze metabolic pathways
- Support plant genomics and comparative genomics research
Website: https://www.arabidopsis.org/
Β