Related Experiment Video
Updated: Feb 28, 2026

Rare Event Detection Using Error-corrected DNA and RNA Sequencing
Published on: August 3, 2018
Revealing the inherent design principles of the genetic code via an error correcting code representation
Ayelet Aharon1,2, Pazit Polak1,2, Gur Yaari3,4,5
1Faculty of Engineering, Bar Ilan University, Ramat Gan, Israel.
Abstract:
The genetic code deterministically maps the 64 possible codons to 20 amino acids, as well as to "START" and "STOP" signals. This universal codon-amino acid mapping (C-AAM) is conserved across almost all living species. The inherent redundancy, arising from mapping 64 codons to only 22 outputs, grants resilience against certain nucleotide substitutions, a property that is conceptually analogous to an error-correcting code (ECC) used in communication systems. ECCs introduce redundancy to protect information against the most probable or most consequential errors during transmission. While coding theory has historically been explored to study the genetic code, biological analogies to traditional communication system elements, such as "source," "encoder," and "channel", remain elusive due to their complexity and partially unknown characteristics. In this study, we adopt the perspective of a communication engineer tasked with reverse-engineering a communication system in which, among its components, only the decoder (the genetic code) is known. By applying this reverse-engineering approach, we introduce the Finding Error Hierarchy (FEH) algorithm, which enables the inference of a hierarchy of nucleotide substitutions against which the genetic code is particularly robust. The methodology also identifies specific amino acid properties that the genetic code preferentially preserves. These findings are validated by their consistency with results from previous studies of the genetic code, conducted using diverse methodologies. We examined mutation patterns at the codon level, allowing consideration of up to three nucleotide substitutions within a single mutation pattern. Since the vast majority of previous studies explored point mutations, the currently derived mutation hierarchy is more comprehensive. This extended hierarchy underscores the biological importance of specific mutations, and offers new perspectives on the functional principles underlying the genetic code.
More Related Videos
11:47Residue-specific Incorporation of Noncanonical Amino Acids into Model Proteins Using an Escherichia coli Cell-free Transcription-translation System
Published on: August 1, 2016
14:02Optimizing the Genetic Incorporation of Chemical Probes into GPCRs for Photo-crosslinking Mapping and Bioorthogonal Chemistry in Live Mammalian Cells
Published on: April 9, 2018
Related Concept Videos
From DNA to Protein
The Central Dogma
The Central Dogma
RNA is the Missing Link Between DNA and Proteins
In the early 1900s, scientists discovered that DNA stores all the information needed for cellular functions and that proteins perform most of these functions. However, the mechanisms of converting genetic information into functional proteins remained unknown for many years. Initially, it was believed that a single gene is...
Genome Copying Errors
Mismatch Repair
Mismatch Repair
The Mutator Protein Family Plays a Key Role in DNA Mismatch Repair
The human genome has more than 3 billion base pairs of DNA per cell. Prior to cell division, that vast amount of genetic...