Related Experiment Video
Updated: Jun 28, 2026

Using the E1A Minigene Tool to Study mRNA Splicing Changes
Published on: April 22, 2021
Origination of the split structure of spliceosomal genes from random genetic sequences
Rahul Regulapati1, Ashwini Bhasi, Chandan Kumar Singh
1Department of Biotechnology, Indian Institute of Technology Madras, Chennai, India.
Abstract:
The mechanism by which protein-coding portions of eukaryotic genes came to be separated by long non-coding stretches of DNA, and the purpose for this perplexing arrangement, have remained unresolved fundamental biological problems for three decades. We report here a plausible solution to this problem based on analysis of open reading frame (ORF) length constraints in the genomes of nine diverse species. If primordial nucleic acid sequences were random in sequence, functional proteins that are innately long would not be encoded due to the frequent occurrence of stop codons. The best possible way that a long protein-coding sequence could have been derived was by evolving a split-structure from the random DNA (or RNA) sequence. Results of the systematic analyses of nine complete genome sequences presented here suggests that perhaps the major underlying structural features of split-genes have evolved due to the indigenous occurrence of split protein-coding genes in primordial random nucleotide sequence. The results also suggest that intron-rich genes containing short exons may have been the original form of genes intrinsically occurring in random DNA, and that intron-poor genes containing long exons were perhaps derived from the original intron-rich genes.
Related Concept Videos
RNA Splicing
RNA Splicing
Exon Recombination
Exon shuffling follows “splice frame rules.” Each exon has three reading...
Gene Conversion
Gene Conversion
Alternative RNA Splicing
There are five types of alternative RNA splicing that vary in the ways the pre-mRNA segments are removed or retained in the mature mRNA. The first...
