Related Experiment Video
Updated: Jul 12, 2026

09:51
Investigating Protein Sequence-structure-dynamics Relationships with Bio3D-web
Published on: July 16, 2017
The PDB data uniformity project.
1Biotechnology Division, National Institute of Standards and Technology, Gaithersburg, MD 20899-8310, USA.
Nucleic Acids Research
|January 11, 2000
Summary
The Protein Data Bank (PDB) is a global archive for macromolecular structures. A new project aims to improve the consistency and reliability of the PDB data for researchers worldwide.
Area of Science:
- Structural biology
- Bioinformatics
- Data science
Background:
- The Protein Data Bank (PDB) serves as the primary global repository for macromolecular structural data.
- Inconsistencies within PDB data can hinder accurate analysis and reproducibility.
Purpose of the Study:
- To describe the ongoing data uniformity project for the Protein Data Bank (PDB).
- To address and rectify data inconsistencies within the PDB archive.
Main Methods:
- Implementation of a data uniformity project.
- Systematic review and standardization of PDB entries.
Main Results:
- The project is actively working to improve data consistency.
- Enhanced data reliability is expected upon project completion.
Conclusions:
- Standardizing PDB data is crucial for advancing structural biology research.
- The data uniformity project will enhance the PDB's utility as a reliable scientific resource.
Related Concept Videos
Gene Families
Gene families consist of groups of genes proposed to have originated from a common ancestor. Typically these arise through events in which a gene or genes are mistakenly duplicated during cell division. Unlike their parent genes (which are subject to selection pressure to maintain function), these gene copies do not need to preserve their sequences and may evolve at a relatively faster rate.
Occasionally these regions can be adapted to take on new roles within the organism, becoming novel genes...
Occasionally these regions can be adapted to take on new roles within the organism, becoming novel genes...
Protein Families
Protein families are groups of homologous proteins; that is, they have similarities in amino acid sequences and three-dimensional structures. Protein families usually occur because of gene duplication, where an additional copy of a gene is inserted into the genome of an organism. Mutations that change the amino acids but still allow the protein to be properly synthesized, will lead to new protein family members. If these new proteins contain similar amino acids in key locations, protein...

