Related Experiment Video
Updated: Jun 28, 2025

Processing the Loblolly Pine PtGen2 cDNA Microarray
Published on: March 20, 2009
Proteogenomic Gene Structure Validation in the Pineapple Genome
Norazrin Ariffin1,2, David Wells Newman1, Michael G Nelson1
1School of Biological Sciences, Faculty of Biology Medicine and Health, MAHSC, University of Manchester, Michael Smith Building, Oxford Road, Manchester M13 9PT, United Kingdom.
Abstract:
MD2 pineapple (Ananas comosus) is the second most important tropical crop that preserves crassulacean acid metabolism (CAM), which has high water-use efficiency and is fast becoming the most consumed fresh fruit worldwide. Despite the significance of environmental efficiency and popularity, until very recently, its genome sequence has not been determined and a high-quality annotated proteome has not been available. Here, we have undertaken a pilot proteogenomic study, analyzing the proteome of MD2 pineapple leaves using liquid chromatography-mass spectrometry (LC-MS/MS), which validates 1781 predicted proteins in the annotated F153 (V3) genome. In addition, a further 603 peptide identifications are found that map exclusively to an independent MD2 transcriptome-derived database but are not found in the standard F153 (V3) annotated proteome. Peptide identifications derived from these MD2 transcripts are also cross-referenced to a more recent and complete MD2 genome annotation, resulting in 402 nonoverlapping peptides, which in turn support 30 high-quality gene candidates novel to both pineapple genomes. Many of the validated F153 (V3) genes are also supported by an independent proteomics data set collected for an ornamental pineapple variety. The contigs and peptides have been mapped to the current F153 genome build and are available as bed files to display a custom gene track on the Ensembl Plants region viewer. These analyses add to the knowledge of experimentally validated pineapple genes and demonstrate the utility of transcript-derived proteomics to discover both novel genes and genetic structure in a plant genome, adding value to its annotation.
Related Concept Videos
Genomic DNA in Eukaryotes
Structure of a Gene
However, only 1% of the DNA is composed of genes that encode proteins; the rest, 99% is non-coding DNA. This non-coding DNA performs...
Genomic DNA in Prokaryotes
Genomic Diversity in Bacteria
Although bacterial genomes are much...
Nucleic Acid Structure
DNA Structure
DNA...
Genome Annotation and Assembly
Chromosome Structure

