"Good annotation practice" for chemical data in biology.

Kirill Degtyarenko1, Marcus Ennis, John S Garavelli

  • 1European Bioinformatics Institute, Wellcome Trust Genome Campus, Hinxton, Cambridge, United Kingdom. kirill@ebi.ac.uk

In Silico Biology
|September 14, 2007
PubMed
Summary

Standardizing chemical language in biological databases is crucial for accurate data representation. This involves clear 2-D diagrams, pronounceable names, and ontologies for context, improving data accessibility and integration.

Related Concept Videos

Genome Annotation and Assembly03:36

Genome Annotation and Assembly

The genome refers to all of the genetic material in an organism. It can range from a few million base pairs in microbial cells to several billion base pairs in many eukaryotic organisms. Genome assembly refers to the process of taking the DNA sequencing data and putting it all back together in a correct order to create a close representation of the original genome. This is followed by the identification of functional elements on the newly assembled genome, a process called genome annotation.
Molecular Models02:00

Molecular Models

Physical models representing molecular architectures of chemical compounds play essential roles in understanding chemistry. The use of molecular models makes it easier to visualize the structures and shapes of atoms and molecules.
Gene Families01:57

Gene Families

Gene families consist of groups of genes proposed to have originated from a common ancestor. Typically these arise through events in which a gene or genes are mistakenly duplicated during cell division. Unlike their parent genes (which are subject to selection pressure to maintain function), these gene copies do not need to preserve their sequences and may evolve at a relatively faster rate.
Occasionally these regions can be adapted to take on new roles within the organism, becoming novel genes...
Chemical Shift: Internal References and Solvent Effects01:17

Chemical Shift: Internal References and Solvent Effects

In an NMR sample, precise measurement of the absolute absorption frequencies of nuclei is difficult. A standard internal reference compound is added, and the frequency difference between the reference signal and sample signals is measured.
The internal reference compound generally used in NMR spectroscopy is tetramethylsilane (TMS). TMS is preferred because it is chemically inert, soluble in NMR solvents, and easily removable. Also, the highly shielded methyl protons in TMS yield an intense...
Chemical Symbols01:09

Chemical Symbols

A chemical symbol is an abbreviation that is used to indicate an element or an atom of an element. For example, the symbol for mercury is Hg. We use the same symbol to indicate one atom of mercury (microscopic domain) or to label a container of many atoms of the element mercury (macroscopic domain).
Some symbols are derived from the common name of the element; others are abbreviations of the name in another language. Most symbols have one or two letters, but three-letter symbols have been used...
EDTA: Chemistry and Properties01:22

EDTA: Chemistry and Properties

Polydentate ligands are most widely used in complexometric titrations because they form more stable complexes with the metal ions than mono- or bidentate ligands due to the chelate effect. Examples of polydentate ligands are ethylenediaminetetraacetic acid (EDTA), crown ethers, and cryptands. The most important feature of optimal polydentate ligands is the ability to form 1:1 complexes in a single-step process. Amino carboxylic acid derivatives are frequently used as complexing agents. EDTA is...