Related Experiment Video
Updated: Apr 17, 2026

Pharmacophore Modeling for Targets with Extensive Ligand Libraries: A Case Study on SARS-CoV-2 Mpro
Published on: September 26, 2025
Multiscale modelling of relationships between protein classes and drug behavior across all diseases using the CANDO
Geetika Sethi, Gaurav Chopra, Ram Samudrala1
1Department of Biomedical Informatics, School of Medicine and Biomedical Sciences, State University of New York (SUNY), 923 Main Street, Buffalo, NY 14203, USA. ram@compbio.org.
Abstract:
We have examined the effect of eight different protein classes (channels, GPCRs, kinases, ligases, nuclear receptors, proteases, phosphatases, transporters) on the benchmarking performance of the CANDO drug discovery and repurposing platform (http://protinfo.org/cando). The first version of the CANDO platform utilizes a matrix of predicted interactions between 48278 proteins and 3733 human ingestible compounds (including FDA approved drugs and supplements) that map to 2030 indications/diseases using a hierarchical chem and bio-informatic fragment based docking with dynamics protocol (> one billion predicted interactions considered). The platform uses similarity of compound-proteome interaction signatures as indicative of similar functional behavior and benchmarking accuracy is calculated across 1439 indications/diseases with more than one approved drug. The CANDO platform yields a significant correlation (0.99, p-value < 0.0001) between the number of proteins considered and benchmarking accuracy obtained indicating the importance of multitargeting for drug discovery. Average benchmarking accuracies range from 6.2 % to 7.6 % for the eight classes when the top 10 ranked compounds are considered, in contrast to a range of 5.5 % to 11.7 % obtained for the comparison/control sets consisting of 10, 100, 1000, and 10000 single best performing proteins. These results are generally two orders of magnitude better than the average accuracy of 0.2% obtained when randomly generated (fully scrambled) matrices are used. Different indications perform well when different classes are used but the best accuracies (up to 11.7% for the top 10 ranked compounds) are achieved when a combination of classes are used containing the broadest distribution of protein folds. Our results illustrate the utility of the CANDO approach and the consideration of different protein classes for devising indication specific protocols for drug repurposing as well as drug discovery.
Insights
The CANDO drug discovery platform shows improved accuracy when considering multiple protein classes and multitargeting. Combining diverse protein classes enhances drug repurposing and discovery protocols for specific indications.
Area of Science:
- Computational drug discovery
- Bioinformatics
- Pharmacology
Background:
- The CANDO platform analyzes drug-target interactions for drug discovery and repurposing.
- Understanding protein class contributions is crucial for optimizing drug discovery protocols.
Purpose of the Study:
- To evaluate the impact of eight distinct protein classes on the CANDO platform's benchmarking performance.
- To assess the role of multitargeting and protein class combinations in drug discovery accuracy.
Main Methods:
- Utilized the CANDO platform, processing over one billion predicted interactions between proteins and compounds.
- Benchmarking accuracy was calculated across 1439 indications with multiple approved drugs.
- Compared performance across eight protein classes against control sets and random matrices.
Main Results:
- A significant positive correlation (0.99) was found between the number of proteins and benchmarking accuracy, highlighting multitargeting's importance.
- Average accuracies ranged from 6.2% to 7.6% for individual protein classes, significantly outperforming random matrices (0.2%).
- Highest accuracies (up to 11.7%) were achieved using a combination of protein classes with diverse protein folds.
Conclusions:
- The CANDO platform demonstrates utility in drug discovery and repurposing.
- Incorporating diverse protein classes and multitargeting strategies enhances prediction accuracy.
- Tailoring protocols based on protein class combinations can optimize indication-specific drug discovery and repurposing.
Related Concept Videos
Structure-Activity Relationships and Drug Design
SAR studies the intricate relationship between a drug's chemical structure and biological activity. It focuses on understanding how modifications to a drug's structure can influence...
Pharmacogenomics: Identification of New Drug Targets
Model Approaches for Pharmacokinetic Data: Distributed Parameter Models
The distributed parameter models are specifically designed to account for variations and differences in some drug classes. This model is particularly useful for assessing regional concentrations of anticancer or...
Drug Discovery: Overview
Pharmacodynamic Models: Additive and Proportional Drug Effect Model
Pharmacodynamic Models: Overview

