Related Experiment Video
Updated: Feb 2, 2026

Quantitative Structure-Activity Relationship, Activity Prediction, and Molecular Dynamics of Non-nucleotide Reverse Transcriptase Inhibitors
Published on: May 9, 2025
Dissecting Machine-Learning Prediction of Molecular Activity: Is an Applicability Domain Needed for Quantitative
Ruifeng Liu1, Hao Wang1, Kyle P Glover2
1Department of Defense, Biotechnology High Performance Computing Software Applications Institute, Telemedicine and Advanced Technology Research Center , U.S. Army Medical Research and Materiel Command , Fort Detrick , Maryland 21702 , United States.
Abstract:
Deep neural networks (DNNs) are the major drivers of recent progress in artificial intelligence. They have emerged as the machine-learning method of choice in solving image and speech recognition problems, and their potential has raised the expectation of similar breakthroughs in other fields of study. In this work, we compared three machine-learning methods-DNN, random forest (a popular conventional method), and variable nearest neighbor (arguably the simplest method)-in their ability to predict the molecular activities of 21 in vivo and in vitro data sets. Surprisingly, the overall performance of the three methods was similar. For molecules with structurally close near neighbors in the training sets, all methods gave reliable predictions, whereas for molecules increasingly dissimilar to the training molecules, all three methods gave progressively poorer predictions. For molecules sharing little to no structural similarity with the training molecules, all three methods gave a nearly constant value-approximately the average activity of all training molecules-as their predictions. The results confirm conclusions deduced from analyzing molecular applicability domains for accurate predictions, i.e., the most important determinant of the accuracy of predicting a molecule is its similarity to the training samples. This highlights the fact that even in the age of deep learning, developing a truly high-quality model relies less on the choice of machine-learning approach and more on the availability of experimental efforts to generate sufficient training data of structurally diverse compounds. The results also indicate that the distance to training molecules offers a natural and intuitive basis for defining applicability domains to flag reliable and unreliable quantitative structure-activity relationship predictions.
Related Concept Videos
Structure-Activity Relationships and Drug Design
SAR studies the intricate relationship between a drug's chemical structure and biological activity. It focuses on understanding how modifications to a drug's structure can influence...
Local Anesthetics: Chemistry and Structure-Activity Relationship
Cholinergic Antagonists: Chemistry and Structure-Activity Relationship
Adrenergic Agonists: Chemistry and Structure-Activity Relationship
Aromatic ring substitutions: Substituting the aromatic ring with –OH groups at positions 3 and 4 yields catecholamines (e.g., epinephrine), which have a high affinity for adrenoceptors. Hydrogen bonding between –OH groups and receptors enhances adrenergic activity.
Separation of...
Predicting Molecular Geometry
Indirect-Acting Cholinergic Agonists: Chemistry and Structure-Activity Relationship
Reversible inhibitors display short to medium durations of action. Short-acting agents include simple alcohols with...

