Related Experiment Video
Updated: Jul 14, 2026

10:41
Identifying Amino Acid Overproducers Using Rare-Codon-Rich Markers
Published on: June 24, 2019
Mining prokaryotic genomes for unknown amino acids: a stop-codon-based approach.
Masashi Fujita1, Hisaaki Mihara, Susumu Goto
1Institute for Chemical Research, Kyoto University, Gokasho, Uji, Kyoto, Japan. fujita@kuicr.kyoto-u.ac.jp <fujita@kuicr.kyoto-u.ac.jp>
BMC Bioinformatics
|June 29, 2007
Summary
A search for a 23rd genetically encoded amino acid found no evidence in 191 prokaryotic genomes. This study confirms that selenocysteine and pyrrolysine remain the only known amino acids encoded by stop codons.
Area of Science:
- Biochemistry
- Genomics
- Molecular Biology
Background:
- Selenocysteine and pyrrolysine are the 21st and 22nd amino acids, encoded by stop codons.
- Previous tRNA gene prediction studies suggested no unknown amino acid, but were inconclusive due to potential atypical tRNA structures.
- A protein-level investigation is needed for independent insight into novel amino acids.
Purpose of the Study:
- To systematically search for a potential 23rd genetically encoded amino acid in prokaryotic genomes.
- To independently validate findings from tRNA gene prediction studies.
- To identify novel selenoproteins and other readthrough proteins.
Main Methods:
- Computational prediction of proteins containing stop-codon-encoded amino acids from 191 prokaryotic genomes.
- Analysis based on conservation patterns of primary amino acid sequences.
- Verification of known selenoproteins and pyrrolysine proteins.
Main Results:
- No candidate for a 23rd amino acid encoded by stop codons was detected.
- The prediction method successfully identified known selenoproteins and pyrrolysine proteins.
- One novel selenoprotein was identified.
Conclusions:
- The findings suggest that a widespread 23rd amino acid encoded by stop codons does not exist.
- The phylogenetic distribution of any potential novel amino acid is likely very limited.
- The developed method can be applied to explore new readthrough events in expanding genomic databases.
Related Concept Videos
Leaky Scanning
During most eukaryotic translation processes, the small 40S ribosome subunit scans an mRNA from its 5' end until it encounters the first start AUG codon. The large 60S ribosomal subunit then joins the smaller one to initiate protein synthesis. The location of the translation initiation is largely determined by the nucleotides near the start codon as there may be multiple translation initiation sites present on the mRNA. Marilyn Kozak discovered that the sequence RCCAUGG (where R stands for...
From DNA to Protein
The flow of genetic information in cells from DNA to mRNA to protein is described by the central dogma, which states that genes specify the sequence of mRNAs, which in turn specify the sequence of amino acids making up all proteins. The decoding of one molecule to another is performed by specific proteins and RNAs. Because the information stored in DNA is so central to cellular function, it makes intuitive sense that the cell would make mRNA copies of this information for protein synthesis...
Translation in Prokaryotes
Prokaryote translation is a complex, highly coordinated process that converts genetic information from mRNA into functional proteins. It involves three stages: initiation, elongation, and termination, each facilitated by specific molecular components.Initiation of TranslationThe process begins with the assembly of the ribosomal subunits and initiation factors on the mRNA. In bacteria, the 30S ribosomal subunit recognizes the Shine-Dalgarno sequence in the mRNA, a conserved region upstream of...
The Central Dogma
Overview
The Central Dogma
The central dogma explains the flow of genetic information from DNA nucleotides to the amino acid sequence of proteins.
RNA is the Missing Link Between DNA and Proteins
In the early 1900s, scientists discovered that DNA stores all the information needed for cellular functions and that proteins perform most of these functions. However, the mechanisms of converting genetic information into functional proteins remained unknown for many years. Initially, it was believed that a single gene is...
RNA is the Missing Link Between DNA and Proteins
In the early 1900s, scientists discovered that DNA stores all the information needed for cellular functions and that proteins perform most of these functions. However, the mechanisms of converting genetic information into functional proteins remained unknown for many years. Initially, it was believed that a single gene is...
Ribosome Profiling
Ribosome profiling or ribo-sequencing is a deep sequencing technique that produces a snapshot of active translation in a cell. It selectively sequences the mRNAs protected by ribosomes to get an insight into a cell’s translation landscape at any given point in time.
Applications of ribosome profiling
Ribosome profiling has many applications, including in vivo monitoring of translation inside a particular organ or tissue type and quantifying new protein synthesis levels.
The technique helps...
Applications of ribosome profiling
Ribosome profiling has many applications, including in vivo monitoring of translation inside a particular organ or tissue type and quantifying new protein synthesis levels.
The technique helps...

