Understanding Primary Databases in Bioinformatics
The question asks to identify examples of primary databases from a given list. Primary databases are fundamental archives in bioinformatics that store and organize experimental data submitted directly from researchers. They serve as the first point of deposition for raw biological data.
Analyzing the Database Examples
Let's examine each option provided:
- (A) GenBank: This is a public, annotated collection of all publicly available DNA sequences. It is maintained by the National Center for Biotechnology Information (NCBI) and is a classic example of a primary nucleotide sequence database.
- (B) EMBL: The European Molecular Biology Laboratory (EMBL) database is another major international nucleotide sequence database, collaborating with GenBank and DDBJ. It also functions as a primary nucleotide sequence database, archiving sequences submitted by researchers in Europe and elsewhere.
- (C) DDBJ: The DNA Data Bank of Japan (DDBJ) is the third major international nucleotide sequence database, working in partnership with GenBank and EMBL. It serves the same purpose as a primary nucleotide sequence database for data originating from Japan and the Asia-Pacific region.
- (D) PDBC (Protein Data Bank): Commonly known as the PDB, this database is the single global repository for the 3D structural data of large biological molecules, such as proteins and nucleic acids. It archives experimental data derived from X-ray crystallography, NMR spectroscopy, and cryo-electron microscopy. The PDB is considered a primary structural database.
Conclusion on Primary Database Classification
Based on the definitions and functions:
- GenBank is a primary database.
- EMBL is a primary database.
- DDBJ is a primary database.
- PDBC (Protein Data Bank) is a primary database.
Therefore, all the listed examples – GenBank (A), EMBL (B), DDBJ (C), and PDBC (D) – are considered primary databases in the field of bioinformatics. The option that includes all of them is the correct one.