Related Experiment Video
Updated: Oct 3, 2025

Phosphoproteomic Strategy for Profiling Osmotic Stress Signaling in Arabidopsis
Published on: June 25, 2020
Predicting protein phosphorylation sites in soybean using interpretable deep tabular learning network
Elham Khalili1, Shahin Ramazi2, Faezeh Ghanati1
1Department of Plant Science, Faculty of Science, Tarbiat Modarres University, Tehran, Iran.
Abstract:
Phosphorylation of proteins is one of the most significant post-translational modifications (PTMs) and plays a crucial role in plant functionality due to its impact on signaling, gene expression, enzyme kinetics, protein stability and interactions. Accurate prediction of plant phosphorylation sites (p-sites) is vital as abnormal regulation of phosphorylation usually leads to plant diseases. However, current experimental methods for PTM prediction suffers from high-computational cost and are error-prone. The present study develops machine learning-based prediction techniques, including a high-performance interpretable deep tabular learning network (TabNet) to improve the prediction of protein p-sites in soybean. Moreover, we use a hybrid feature set of sequential-based features, physicochemical properties and position-specific scoring matrices to predict serine (Ser/S), threonine (Thr/T) and tyrosine (Tyr/Y) p-sites in soybean for the first time. The experimentally verified p-sites data of soybean proteins are collected from the eukaryotic phosphorylation sites database and database post-translational modification. We then remove the redundant set of positive and negative samples by dropping protein sequences with >40% similarity. It is found that the developed techniques perform >70% in terms of accuracy. The results demonstrate that the TabNet model is the best performing classifier using hybrid features and with window size of 13, resulted in 78.96 and 77.24% sensitivity and specificity, respectively. The results indicate that the TabNet method has advantages in terms of high-performance and interpretability. The proposed technique can automatically analyze the data without any measurement errors and any human intervention. Furthermore, it can be used to predict putative protein p-sites in plants effectively. The collected dataset and source code are publicly deposited at https://github.com/Elham-khalili/Soybean-P-sites-Prediction.
Related Concept Videos
Protein-protein Interfaces
Protein Networks
These interactions can be represented through maps depicting protein-protein interaction networks, represented as nodes and edges. Nodes are circles that are representative of a protein,...
Protein Kinases and Phosphatases
Protein kinases
Many proteins in the cell are regulated by phosphorylation, the addition of a phosphate group. A family of enzymes called kinases...
Conserved Binding Sites
Binding sites are often located in large pockets, and if their location on a protein’s surface is unknown, it can be predicted using various approaches. The energetic method computationally...
Phosphorylation
During phosphorylation, protein kinases transfer the terminal phosphate group of ATP to specific amino acid side chains of substrate proteins. Serine, threonine, and tyrosine are the most commonly...
PI3K/mTOR/AKT Signaling Pathway

