A sequence-based hybrid predictor for identifying conformationally ambivalent regions in proteins

Yu-Cheng Liu1, Meng-Han Yang, Win-Li Lin

  • 1Institute of Biomedical Engineering, National Taiwan University, Taipei, Taiwan, Republic of China. f90548051@ntu.edu.tw

BMC Genomics
|December 5, 2009
PubMed
Summary

This study introduces a novel sequence-based predictor for identifying flexible protein regions. The hybrid predictor combines two supervised learning algorithms, outperforming existing methods and offering valuable insights for biologists.

Related Concept Videos

Conserved Binding Sites01:49

Conserved Binding Sites

Many proteins’ biological role depends on their interactions with their ligands, small molecules that bind to specific locations on the protein known as ligand-binding sites. Ligand-binding sites are often conserved among homologous proteins as these sites are critical for protein function.
Binding sites are often located in large pockets, and if their location on a protein’s surface is unknown, it can be predicted using various approaches. The energetic method computationally analyses the...
Conservation of Protein Domains Over Different Proteins02:26

Conservation of Protein Domains Over Different Proteins

Protein domains are small structurally independent units that are part of a single amino acid chain.  Although these domains are often structurally independent, they may rely on synergistic effects to perform their functions as part of a larger protein. Protein domains may be conserved within the same organism, as well as across different organisms.
A limited set of protein domains often duplicate and recombine during evolution. These domains can be organized in different combinations to form...
Conservation of Protein Domains02:26

Conservation of Protein Domains

Protein domains are small structurally independent units that are part of a single amino acid chain.  Although these domains are often structurally independent, they may rely on synergistic effects to perform their functions as part of a larger protein. Protein domains may be conserved within the same organism, as well as across different organisms.
A limited set of protein domains often duplicate and recombine during evolution. These domains can be organized in different combinations to form...
Multi-species Conserved Sequences02:51

Multi-species Conserved Sequences

Next-generation sequencing technologies have created large genomic databases of a variety of animals and plants. Ever since the human genome project was completed, scientists studied the genome of primates, mammals, and other phylogenetically distant living beings. Such large-scale  studies have provided new insights into the evolutionary relationship between organisms.
Although the genome of each species varies greatly from each other, a few sequences are highly conserved. Such conserved DNA...
Mitochondrial Precursor Proteins01:39

Mitochondrial Precursor Proteins

Mitochondrial precursors are partially unfolded or loosely folded polypeptide chains. Newly synthesized precursors are inhibited from spontaneously folding into their native conformation by the cytosolic chaperones, heat shock proteins 70 (Hsp70), and mitochondrial import stimulation factors (MSFs). Precursors bound to MSFs are guided to the TOM70-TOM37 receptors, while precursors bound to Hsp70  chaperones are targetted to TOM20-TOM22 receptor complexes.
Most of the mitochondrial precursors...
Peptide Identification Using Tandem Mass Spectrometry01:33

Peptide Identification Using Tandem Mass Spectrometry

Tandem mass spectrometry, also known as MS/MS or MS2, is an analytical technique that employs two mass analyzers. Essentially it is a series of mass spectrometers that helps isolate a particular biomolecule and then helps study its chemical properties.
This technique helps gather information regarding the protein from which the peptide was obtained and to study the peptides’ amino acid sequence. Identifying peptides from a complex mixture is an important component of the growing field of...