Related Experiment Video
Updated: Mar 30, 2026

Highly Sensitive and Quantitative Detection of Proteins and Their Isoforms by Capillary Isoelectric Focusing Method
Published on: September 19, 2018
Accurate estimation of isoelectric point of protein and peptide based on amino acid sequences
Enrique Audain1, Yassel Ramos2, Henning Hermjakob3
1Department of Proteomics, Center of Molecular Immunology.
Motivation:
In any macromolecular polyprotic system-for example protein, DNA or RNA-the isoelectric point-commonly referred to as the pI-can be defined as the point of singularity in a titration curve, corresponding to the solution pH value at which the net overall surface charge-and thus the electrophoretic mobility-of the ampholyte sums to zero. Different modern analytical biochemistry and proteomics methods depend on the isoelectric point as a principal feature for protein and peptide characterization. Protein separation by isoelectric point is a critical part of 2-D gel electrophoresis, a key precursor of proteomics, where discrete spots can be digested in-gel, and proteins subsequently identified by analytical mass spectrometry. Peptide fractionation according to their pI is also widely used in current proteomics sample preparation procedures previous to the LC-MS/MS analysis. Therefore accurate theoretical prediction of pI would expedite such analysis. While such pI calculation is widely used, it remains largely untested, motivating our efforts to benchmark pI prediction methods.
Results:
Using data from the database PIP-DB and one publically available dataset as our reference gold standard, we have undertaken the benchmarking of pI calculation methods. We find that methods vary in their accuracy and are highly sensitive to the choice of basis set. The machine-learning algorithms, especially the SVM-based algorithm, showed a superior performance when studying peptide mixtures. In general, learning-based pI prediction methods (such as Cofactor, SVM and Branca) require a large training dataset and their resulting performance will strongly depend of the quality of that data. In contrast with Iterative methods, machine-learning algorithms have the advantage of being able to add new features to improve the accuracy of prediction.
Contact:
yperez@ebi.ac.uk
Availability And Implementation:
The software and data are freely available at https://github.com/ypriverol/pIRSupplementary information: Supplementary data are available at Bioinformatics online.
Related Concept Videos
Two-dimensional Gel Electrophoresis
The first dimension separation uses the isoelectric focusing or IEF technique performed on immobilized pH gradient (IPG) strips that separate proteins according to their isoelectric points.
Biological samples, such...
SDS-PAGE
A variation of gel electrophoresis, termed polyacrylamide gel electrophoresis (PAGE), is commonly used for separating proteins according to their molecular size by passing them through a polyacrylamide gel. Because of the varying charges associated with amino acid side chains, PAGE can be used to separate intact...
What are Proteins?

