Related Experiment Video
Updated: May 3, 2026

14:37
Modeling an Enzyme Active Site using Molecular Visualization Freeware
Published on: December 25, 2021
11.6K
Data model, dictionaries, and desiderata for biomolecular simulation data indexing and sharing
Julien C Thibault, Daniel R Roe, Julio C Facelli1
1Department of Biomedical Informatics, University of Utah, Salt Lake City, UT, USA. julio.facelli@utah.edu.
Journal of Cheminformatics
|February 4, 2014
Summary
A new logical model and data dictionaries standardize biomolecular simulation data, improving data sharing and reuse for computational chemistry and molecular dynamics research. This facilitates better data management and exploration in the field.
Area of Science:
- Computational Biology
- Biomolecular Simulations
- Data Science
Background:
- Limited environments exist for sharing biomolecular simulation data and fostering collaborative exploration.
- Increasing data volume and complexity necessitate improved tools for managing and sharing simulation outputs.
- Growing application of simulation methods highlights the need for standardized data representation.
Purpose of the Study:
- To assess community needs for data representation standards in biomolecular simulations.
- To guide the development of future repositories for biomolecular simulation data.
- To establish a common logical model for representing diverse simulation data.
Main Methods:
- Community feedback gathered through surveys and interviews informed the selection of common data elements.
- Data elements were integrated to encompass concepts from quantum chemistry and molecular dynamics.
- A logical model and associated dictionaries were developed for database and API design.
Main Results:
- A comprehensive list of common data elements for biomolecular simulations was identified and refined.
- A logical model was created to structure data for new databases and application programming interfaces.
- SQL queryable and Java API-accessible dictionaries were implemented using the Apache Lucene search engine.
Conclusions:
- The proposed model and dictionaries offer a robust yet simple representation for biomolecular simulation concepts.
- This framework is extensible and supports future development of terminologies and ontologies.
- Use cases demonstrate benefits for data storage, indexing, and presentation, enhancing data accessibility and reuse.
Related Concept Videos
Molecular Models
37.5K
Physical models representing molecular architectures of chemical compounds play essential roles in understanding chemistry. The use of molecular models makes it easier to visualize the structures and shapes of atoms and molecules.
37.5K
Gene Families
8.0K
Gene families consist of groups of genes proposed to have originated from a common ancestor. Typically these arise through events in which a gene or genes are mistakenly duplicated during cell division. Unlike their parent genes (which are subject to selection pressure to maintain function), these gene copies do not need to preserve their sequences and may evolve at a relatively faster rate.
Occasionally these regions can be adapted to take on new roles within the organism, becoming novel genes...
Occasionally these regions can be adapted to take on new roles within the organism, becoming novel genes...
8.0K
Globular and Fibrous Proteins
34.0K
Many proteins can be classified into two distinct subtypes - globular or fibrous. These two types differ in their shapes and solubilities.
Globular proteins are also known as spheroproteins and typically are approximately round in shape. They contain a mix of amino acid types and contain differing sequences in their primary structures. Globular proteins have many different functions, such as enzymes, cellular messengers, and molecular transporters. These roles often require the proteins to be...
Globular proteins are also known as spheroproteins and typically are approximately round in shape. They contain a mix of amino acid types and contain differing sequences in their primary structures. Globular proteins have many different functions, such as enzymes, cellular messengers, and molecular transporters. These roles often require the proteins to be...
34.0K
Protein Families
13.3K
Protein families are groups of homologous proteins; that is, they have similarities in amino acid sequences and three-dimensional structures. Protein families usually occur because of gene duplication, where an additional copy of a gene is inserted into the genome of an organism. Mutations that change the amino acids but still allow the protein to be properly synthesized, will lead to new protein family members. If these new proteins contain similar amino acids in key...
13.3K
X-ray Diffraction of Biological Samples
3.8K
X-ray diffraction or XRD is an analytical tool that utilizes X-rays to study ordered structures such as crystalline organic and inorganic samples, polycrystalline materials, proteins, carbohydrates, and drugs.
According to Bragg's law, when X-rays strike the sample positioned on a stage, the rays are scattered by the electron clouds around the sample atoms. The X-ray diffraction or scattering is caused by constructive interference of the X-ray waves that reflect off the internal...
According to Bragg's law, when X-rays strike the sample positioned on a stage, the rays are scattered by the electron clouds around the sample atoms. The X-ray diffraction or scattering is caused by constructive interference of the X-ray waves that reflect off the internal...
3.8K

