Related Experiment Video
Updated: May 13, 2026

A Psychophysics Paradigm for the Collection and Analysis of Similarity Judgments
Published on: March 1, 2022
QSAR with experimental and predictive distributions: an information theoretic approach for assessing model quality
David J Wood1, Lars Carlsson, Martin Eklund
1Novartis, Horsham, UK. davejwood@gmail.com
Quantitative structure-activity relationship (QSAR) predictions are proposed as probability distributions. This approach, using Kullback-Leibler divergence, assesses model quality and identifies algorithms for accurate compound-specific predictions.
Area of Science:
- Computational chemistry
- Drug discovery
- Machine learning
Background:
- Quantitative structure-activity relationship (QSAR) models are crucial for predicting drug properties.
- Current QSAR predictions often lack explicit uncertainty quantification.
- Accurate uncertainty estimation is vital for reliable drug development decisions.
Purpose of the Study:
- To represent QSAR predictions as probability distributions for robust uncertainty quantification.
- To evaluate machine learning algorithms and error estimation methods for generating predictive distributions.
- To assess the utility of predictive distributions in estimating compound property profiles.
Main Methods:
- Utilized Kullback-Leibler (KL) divergence to measure the quality of predictive distributions against experimental data.
- Assessed various machine learning algorithms and error estimation techniques on AstraZeneca's DMPK datasets.
- Developed reliability indices to associate predictive distribution tightness with prediction reliability.
Main Results:
- Identified specific algorithm and error estimation combinations that yield accurate and valid compound-specific predictive distributions.
- Demonstrated that KL divergence effectively quantifies the quality of QSAR predictive distributions.
- Showcased the use of reliability indices to differentiate between high- and low-confidence predictions.
Conclusions:
- Representing QSAR predictions as probability distributions enhances model interpretability and reliability.
- The KL-divergence framework provides a robust method for evaluating predictive distribution quality.
- Validated predictive distributions can inform the probability of a compound meeting specific target profiles, aiding in compound selection.
Related Concept Videos
Model Approaches for Pharmacokinetic Data: Distributed Parameter Models
The distributed parameter models are specifically designed to account for variations and differences in some drug classes. This model is particularly useful for assessing regional concentrations of anticancer or...
Probability Distributions
A discrete probability distribution is a probability distribution of discrete random variables. It can be categorized into binomial probability distribution and Poisson probability...
Analysis Methods of Pharmacokinetic Data: Model and Model-Independent Approaches
The model approach uses mathematical models to describe changes in drug concentration over time. Pharmacokinetic models help characterize drug behavior in patients, predict drug concentration in the body fluids, calculate optimum dosage regimens, and evaluate the risk of toxicity. However, ensuring that the model fits the experimental data accurately...
Pharmacokinetic Models: Comparison and Selection Criterion
Physiological models take a detailed approach by considering specific molecular processes. They can predict drug distribution, metabolism, and elimination changes, providing a comprehensive understanding of how drugs interact with the body.
Model Approaches for Pharmacokinetic Data: Compartment Models
Two primary types of compartment models are recognized: mammillary and catenary. The more...
Mechanistic Models: Compartment Models in Individual and Population Analysis