Related Experiment Video
Updated: Jun 24, 2026

08:23
De novo Identification of Actively Translated Open Reading Frames with Ribosome Profiling Data
Published on: February 18, 2022
Obtaining accurate translations from expressed sequence tags
1Institute of Evolutionary Biology, University of Edinburgh, Edinburgh, UK.
Methods in Molecular Biology (Clifton, N.J.)
|March 12, 2009
Summary
Accurate polypeptide translations are crucial for analyzing expressed sequence tags (ESTs). Our prot4EST pipeline improves EST annotation accuracy without needing extensive training data.
Area of Science:
- Genomics
- Bioinformatics
Background:
- Expressed sequence tags (ESTs) are vital for genome investigation but suffer from errors and incomplete transcripts.
- Accurate polypeptide translations are essential for effective EST annotation, yet current methods often require substantial training data unavailable for many species.
- Neglected genomes present unique challenges due to limited full-length gene sequence resources.
Purpose of the Study:
- To develop an improved method for polypeptide translation from expressed sequence tags (ESTs).
- To overcome the limitations of existing EST translation tools, particularly the need for large training datasets.
- To enhance the accuracy of downstream annotation for ESTs from understudied genomes.
Main Methods:
- Development of the prot4EST polypeptide prediction pipeline.
- Integration of freely available software tools into a cohesive workflow.
- Application of the pipeline to EST data from neglected genomes.
Main Results:
- The prot4EST pipeline generates more accurate polypeptide translations compared to single-method approaches.
- The integrated pipeline effectively addresses the deficit in training data for EST translation.
- Improved annotation accuracy is achieved for ESTs, facilitating genomic research.
Conclusions:
- The prot4EST pipeline offers a robust solution for accurate EST translation, especially for species with limited genomic resources.
- This approach significantly enhances the utility of EST data for genomic and transcriptomic studies.
- Prot4EST represents a valuable advancement in bioinformatics tools for genome annotation.
Related Concept Videos
Ribosome Profiling
Ribosome profiling or ribo-sequencing is a deep sequencing technique that produces a snapshot of active translation in a cell. It selectively sequences the mRNAs protected by ribosomes to get an insight into a cell’s translation landscape at any given point in time.
Applications of ribosome profiling
Ribosome profiling has many applications, including in vivo monitoring of translation inside a particular organ or tissue type and quantifying new protein synthesis levels.
The technique helps...
Applications of ribosome profiling
Ribosome profiling has many applications, including in vivo monitoring of translation inside a particular organ or tissue type and quantifying new protein synthesis levels.
The technique helps...
RNA-seq
RNA sequencing, or RNA-Seq, is a high-throughput sequencing technology used to study the transcriptome of a cell. Transcriptomics helps to interpret the functional elements of a genome and identify the molecular constituents of an organism. Additionally, it also helps in understanding the development of an organism and the occurrence of diseases.
Before the discovery of RNA-seq, microarray-based methods and Sanger sequencing were used for transcriptome analysis. However, while microarray-based...
Before the discovery of RNA-seq, microarray-based methods and Sanger sequencing were used for transcriptome analysis. However, while microarray-based...
Improving Translational Accuracy
Base complementarity between the three base pairs of mRNA codon and the tRNA anticodon is not a failsafe mechanism. Inaccuracies can range from a single mismatch to no correct base pairing at all. The free energy difference between the correct and nearly correct base pairs can be as small as 3 kcal/ mol. With complementarity being the only proofreading step, the estimated error frequency would be one wrong amino acid in every 100 amino acids incorporated. However, error frequencies observed in...
Improving Translational Accuracy
Base complementarity between the three base pairs of mRNA codon and the tRNA anticodon is not a failsafe mechanism. Inaccuracies can range from a single mismatch to no correct base pairing at all. The free energy difference between the correct and nearly correct base pairs can be as small as 3 kcal/ mol. With complementarity being the only proofreading step, the estimated error frequency would be one wrong amino acid in every 100 amino acids incorporated. However, error frequencies observed in...
Leaky Scanning
During most eukaryotic translation processes, the small 40S ribosome subunit scans an mRNA from its 5' end until it encounters the first start AUG codon. The large 60S ribosomal subunit then joins the smaller one to initiate protein synthesis. The location of the translation initiation is largely determined by the nucleotides near the start codon as there may be multiple translation initiation sites present on the mRNA. Marilyn Kozak discovered that the sequence RCCAUGG (where R stands for...
