Related Experiment Video
Updated: Jan 22, 2026

13:42
RNA Secondary Structure Prediction Using High-throughput SHAPE
Published on: May 31, 2013
32.2K
ENTRNA: a framework to predict RNA foldability
Congzhe Su1, Jeffery D Weir2, Fei Zhang3
1School of Computing, Informatics, Decision Systems Engineering, Arizona State University, Tempe, AZ, 85281, USA.
BMC Bioinformatics
|July 5, 2019
Summary
We introduce ENTRNA, a data-driven framework for predicting RNA foldability. This approach enhances RNA design by assessing sequence-structure pair likelihood, improving upon traditional methods.
Area of Science:
- Computational Biology
- Bioinformatics
- Molecular Biology
Background:
- RNA molecules are vital for cellular functions, with their spatial structures dictating biological roles.
- Understanding RNA folding, especially secondary structures, is crucial for deciphering these functions.
- Current RNA design methods often treat it as a structure prediction or inverse folding problem, heavily relying on free energy calculations.
Purpose of the Study:
- To develop a novel data-driven framework, ENTRNA, for predicting RNA foldability.
- To address limitations in existing RNA design methodologies by reframing it as a foldability prediction problem.
- To introduce a new feature, Sequence Segment Entropy (SSE), for measuring RNA sequence diversity.
Main Methods:
- Developed the ENTRNA framework using a Positive-Unlabeled learning approach.
- Incorporated Sequence Segment Entropy (SSE) alongside traditional sequence and structural features.
- Trained and validated ENTRNA on extensive datasets from the RNASTRAND database (1024 pseudoknot-free, 1060 pseudoknotted RNAs).
- Tested model robustness on independent datasets from the PDB (206 pseudoknot-free, 93 pseudoknotted RNAs).
Main Results:
- ENTRNA achieved 86.5% sensitivity on the pseudoknot-free training set and 80.6% on the testing set.
- For pseudoknotted RNAs, ENTRNA demonstrated 81.5% sensitivity on the training set and 71.0% on the testing set.
- Successfully predicted the foldability of 4 out of 5 long, structurally complex synthetic RNAs.
Conclusions:
- Reformulated RNA design as a foldability prediction problem, estimating the likelihood of sequence-structure pair co-existence.
- The ENTRNA framework offers potential advancements for both RNA structure prediction and inverse folding problems.
- This work highlights the utility of data-driven approaches in RNA research.
Related Concept Videos
Predicting Molecular Geometry
45.5K
VSEPR Theory for Determination of Electron Pair Geometries
45.5K
Prediction Intervals
3.3K
The interval estimate of any variable is known as the prediction interval. It helps decide if a point estimate is dependable.
However, the point estimate is most likely not the exact value of the population parameter, but close to it. After calculating point estimates, we construct interval estimates, called confidence intervals or prediction intervals. This prediction interval comprises a range of values unlike the point estimate and is a better predictor of the observed sample value, y.
However, the point estimate is most likely not the exact value of the population parameter, but close to it. After calculating point estimates, we construct interval estimates, called confidence intervals or prediction intervals. This prediction interval comprises a range of values unlike the point estimate and is a better predictor of the observed sample value, y.
3.3K
RNA Stability
35.6K
Intact DNA strands can be found in fossils, while scientists sometimes struggle to keep RNA intact under laboratory conditions. The structural variations between RNA and DNA underlie the differences in their stability and longevity. Because DNA is double-stranded, it is inherently more stable. The single-stranded structure of RNA is less stable but also more flexible and can form weak internal bonds. Additionally, most RNAs in the cell are relatively short, while DNA can be up to 250 million...
35.6K
RNA Interference
27.9K
RNA interference (RNAi) is a process in which a small non-coding RNA molecule blocks the post-transcriptional expression of a gene by binding to its messenger RNA (mRNA) and preventing the protein from being translated.
This process occurs naturally in cells, often through the activity of genomically-encoded microRNAs. Researchers can take advantage of this mechanism by introducing synthetic RNAs to deactivate specific genes for research or therapeutic purposes. For example, RNAi could be used...
This process occurs naturally in cells, often through the activity of genomically-encoded microRNAs. Researchers can take advantage of this mechanism by introducing synthetic RNAs to deactivate specific genes for research or therapeutic purposes. For example, RNAi could be used...
27.9K
RNA Structure
78.9K
Overview
The basic structure of RNA consists of a five-carbon sugar and one of four nitrogenous bases. Although most RNA is single-stranded, it can form complex secondary and tertiary structures. Such structures play essential roles in the regulation of transcription and translation.
Different Types of RNA Have the Same Basic Structure
There are three main types of ribonucleic acid (RNA): messenger RNA (mRNA), transfer RNA (tRNA), and ribosomal RNA (rRNA). All three RNA types consist of a...
The basic structure of RNA consists of a five-carbon sugar and one of four nitrogenous bases. Although most RNA is single-stranded, it can form complex secondary and tertiary structures. Such structures play essential roles in the regulation of transcription and translation.
Different Types of RNA Have the Same Basic Structure
There are three main types of ribonucleic acid (RNA): messenger RNA (mRNA), transfer RNA (tRNA), and ribosomal RNA (rRNA). All three RNA types consist of a...
78.9K
RNA Splicing
60.4K
Splicing is the process by which eukaryotic RNA is edited before its translation into protein. The RNA strand transcribed from eukaryotic DNA is called the primary transcript. The primary transcripts that become mRNAs are called precursor messenger RNAs (pre-mRNAs). Eukaryotic pre-mRNA contains alternating sequences of exons and introns. Exons are nucleotide sequences that code for proteins, whereas introns are the non-coding regions. In RNA splicing, introns are removed and exons are bonded...
60.4K

