Related Experiment Video
Updated: May 13, 2026

G2-seq: A High Throughput Sequencing-based Technique for Identifying Late Replicating Regions of the Genome
Published on: March 22, 2018
The Genomes of Oryza sativa: a history of duplications
1Beijing Institute of Genomics of the Chinese Academy of Sciences, Beijing Genomics Institute, Beijing Proteomics Institute, China. junyu@genomics.org.cn <junyu@genomics.org.cn>
Abstract:
We report improved whole-genome shotgun sequences for the genomes of indica and japonica rice, both with multimegabase contiguity, or almost 1,000-fold improvement over the drafts of 2002. Tested against a nonredundant collection of 19,079 full-length cDNAs, 97.7% of the genes are aligned, without fragmentation, to the mapped super-scaffolds of one or the other genome. We introduce a gene identification procedure for plants that does not rely on similarity to known genes to remove erroneous predictions resulting from transposable elements. Using the available EST data to adjust for residual errors in the predictions, the estimated gene count is at least 38,000-40,000. Only 2%-3% of the genes are unique to any one subspecies, comparable to the amount of sequence that might still be missing. Despite this lack of variation in gene content, there is enormous variation in the intergenic regions. At least a quarter of the two sequences could not be aligned, and where they could be aligned, single nucleotide polymorphism (SNP) rates varied from as little as 3.0 SNP/kb in the coding regions to 27.6 SNP/kb in the transposable elements. A more inclusive new approach for analyzing duplication history is introduced here. It reveals an ancient whole-genome duplication, a recent segmental duplication on Chromosomes 11 and 12, and massive ongoing individual gene duplications. We find 18 distinct pairs of duplicated segments that cover 65.7% of the genome; 17 of these pairs date back to a common time before the divergence of the grasses. More important, ongoing individual gene duplications provide a never-ending source of raw material for gene genesis and are major contributors to the differences between members of the grass family.
More Related Videos
09:32An Array-based Comparative Genomic Hybridization Platform for Efficient Detection of Copy Number Variations in Fast Neutron-induced Medicago truncatula Mutants
Published on: November 8, 2017
07:18Obtaining High-Quality Transcriptome Data from Cereal Seeds by a Modified Method for Gene Expression Profiling
Published on: May 21, 2020
Related Concept Videos
Genome Size and the Evolution of New Genes
Gene Families
Occasionally these regions can be adapted to take on new roles within the organism, becoming novel genes...
Chromosome Structure
The centromere is a DNA sequence that links sister chromatids. This is also where kinetochores, protein complexes to which spindle microtubules attach, are constructed after the chromosome is replicated. The kinetochores allow the spindle microtubules to move the chromosomes within the cell during cell division.
Telomeres consist of non-coding repetitive nucleotide...
Gene Duplication and Divergence
The duplicated copies of the gene are called Paralogs. Paralogs with similar sequences and functions form a gene family. Across several species, a large number of gene families are characterized.
Genome Size and the Evolution of New Genes
Gene Families
Occasionally these regions can be adapted to take on new roles within the organism, becoming novel genes...