Prediction of posttranslational modification sites from amino acid sequences with kernel methods

Yan Xu1, Xiaobo Wang2, Yongcui Wang2

  • 1Department of Information and Computer Science, University of Science and Technology Beijing, Beijing 100083, China.

Summary

A new computational method, position-specific propensity matrices (PSPM), effectively predicts protein post-translational modification (PTM) sites. This tool offers a faster, cost-effective alternative to experimental methods for PTM site identification.

Related Concept Videos

Conserved Binding Sites01:49

Conserved Binding Sites

Many proteins’ biological role depends on their interactions with their ligands, small molecules that bind to specific locations on the protein known as ligand-binding sites. Ligand-binding sites are often conserved among homologous proteins as these sites are critical for protein function.
Binding sites are often located in large pockets, and if their location on a protein’s surface is unknown, it can be predicted using various approaches. The energetic method computationally...
4.1K
Protein Modifications in the RER01:26

Protein Modifications in the RER

Modification of secretory and transmembrane proteins entering the rough ER begins in the ER lumen. These modifications aid in protein folding and stabilize the acquired tertiary structure. Protein modifications in the rough ER co-occur at different stages of protein folding.
Broadly, these modifications can be categorized into four main categories — glycosylation, formation of disulfide bonds, assembly of protein subunits, and specific proteolytic cleavages like removal of signal...
5.7K
Signal Sequences and Sorting Receptors01:41

Signal Sequences and Sorting Receptors

Signal sequences are short amino acid sequences that guide newly synthesized proteins to their proper location within the cell. Classical signal sequences are fifteen to sixty amino acids long and present at the N-terminus of a polypeptide chain. Each signal sequence has a conserved segment of basic residues towards their N terminus, a hydrophobic core, and a C-terminus rich in polar residues. The C-terminus also contains a signal cleavage site and features a -3 -1 sequence motif. The -3-1...
9.9K
Conservation of Protein Domains Over Different Proteins02:26

Conservation of Protein Domains Over Different Proteins

Protein domains are small structurally independent units that are part of a single amino acid chain.  Although these domains are often structurally independent, they may rely on synergistic effects to perform their functions as part of a larger protein. Protein domains may be conserved within the same organism, as well as across different organisms.
A limited set of protein domains often duplicate and recombine during evolution. These domains can be organized in different combinations to...
11.8K
Pre-mRNA Processing: Modification of pre-mRNA Ends01:35

Pre-mRNA Processing: Modification of pre-mRNA Ends

In eukaryotic cells, transcripts made by RNA polymerase are modified and processed before exiting the nucleus. Unprocessed RNA is called precursor mRNA or pre-mRNA to distinguish it from mature mRNA.
Once about 20-40 ribonucleotides have been joined together by RNA polymerase, a group of enzymes adds a cap to the 5' end of the growing transcript. In this process, a 5' phosphate is replaced by modified guanosine that has a methyl group attached (7-methyl guanosine). This 5' cap helps...
14.2K
Proteomics01:33

Proteomics

A proteome is the entire set of proteins that a cell type produces. We can study proteomes using the knowledge of genomes because genes code for mRNAs, and the mRNAs encode proteins. Although mRNA analysis is a step in the right direction, not all mRNAs are translated into proteins.
Proteomics is the study of proteomes' function. It involves the large-scale systematic study of the proteome to denote the protein complement expressed by a genome. Scientist Mark Wilkins coined the term...
7.5K