Bioinformatics predictions of localization and targeting

Shruti Rastogi1, Burkhard Rost

  • 1Department of Biochemistry and Molecular Biophysics, Columbia University and Columbia University Center for Computational Biology and Bioinformatics (C2B2), New York, NY, USA.

Summary

Predicting protein subcellular localization is crucial for understanding protein function. Advanced machine learning methods have significantly improved the accuracy of these predictions, aiding in post-genomic annotation efforts.

Related Concept Videos

Conserved Binding Sites01:49

Conserved Binding Sites

Many proteins’ biological role depends on their interactions with their ligands, small molecules that bind to specific locations on the protein known as ligand-binding sites. Ligand-binding sites are often conserved among homologous proteins as these sites are critical for protein function.
Binding sites are often located in large pockets, and if their location on a protein’s surface is unknown, it can be predicted using various approaches. The energetic method computationally analyses the...
Regulated mRNA Transport02:22

Regulated mRNA Transport

In eukaryotes, transcription and translation are compartmentalized; an mRNA is first synthesized in the nucleus and then selectively transported to the cytoplasm for protein synthesis. Before transport, a pre-mRNA undergoes several steps of post-transcriptional modifications including splicing, 5' capping, and the addition of a poly-adenine tail. Various proteins bind to the pre-mRNA during these modifications. The mRNA transport takes place with the help of multiple proteins playing specific...
Nuclear Localization Signals and Import01:46

Nuclear Localization Signals and Import

Proteins targeted to the nucleus carry short stretches of amino acid sequences called the nuclear localization signal or NLS. Classical nuclear localization signals are of two types: monopartite and bipartite NLS. Monopartite classical NLS (cNLS) consists of a single cluster of 4-8 amino acids. Bipartite cNLS consists of two clusters of  2-3 amino acids and a 9-12 residue long proline-rich linker bridging the two clusters. Signal clusters are rich in positively charged amino acids such as...
Directing Proteins to the Rough Endoplasmic Reticulum01:34

Directing Proteins to the Rough Endoplasmic Reticulum

The organelle-specific signaling sequences direct proteins synthesized in the cytosol to their final destination like ER, mitochondria, peroxisomes, etc. Some of the proteins directed to ER are then trafficked via vesicles to other organelles within the cell or the extracellular environment through the Golgi complex. For example, the rough ER synthesizes soluble proteins for transportation to the lysosomes or secretion out of the cell. It can also synthesize transmembrane proteins that can...
Protein-protein Interfaces02:04

Protein-protein Interfaces

Many proteins form complexes to carry out their functions, making protein-protein interactions (PPIs) essential for an organism's survival. Most PPIs are stabilized by numerous weak noncovalent chemical forces. The physical shape of the interfaces determines the way two proteins interact. Many globular proteins have closely-matching shapes on their surfaces, which form a large number of weak bonds. Additionally, many PPIs occur between two helices or between a surface cleft and a polypeptide...
Signal Sequences and Sorting Receptors01:41

Signal Sequences and Sorting Receptors

Signal sequences are short amino acid sequences that guide newly synthesized proteins to their proper location within the cell. Classical signal sequences are fifteen to sixty amino acids long and present at the N-terminus of a polypeptide chain. Each signal sequence has a conserved segment of basic residues towards their N terminus, a hydrophobic core, and a C-terminus rich in polar residues. The C-terminus also contains a signal cleavage site and features a -3 -1 sequence motif. The -3-1...