Related Experiment Video
Updated: Feb 12, 2026

In Silico Modeling Method for Computational Aquatic Toxicology of Endocrine Disruptors: A Software-Based Approach Using QSAR Toolbox
Published on: August 28, 2019
Making reliable negative predictions of human skin sensitisation using an in silico fragmentation approach
Martyn L Chilton1, Donna S Macmillan1, Thomas Steger-Hartmann2
1Lhasa Limited, Granary Wharf House, 2 Canal Wharf, Leeds, LS11 5PS, UK.
Abstract:
A previously published fragmentation method for making reliable negative in silico predictions has been applied to the problem of predicting skin sensitisation in humans, making use of a dataset of over 2750 chemicals with publicly available skin sensitisation data from 18 in vivo assays. An assay hierarchy was designed to enable the classification of chemicals within this dataset as either sensitisers or non-sensitisers where data from more than one in vivo test was available. The negative prediction approach was validated internally, using a 5-fold cross-validation, and externally, against a proprietary dataset of approximately 1000 chemicals with in vivo reference data shared by members of the pharmaceutical, nutritional, and personal care industries. The negative predictivity for this proprietary dataset was high in all cases (>75%), and the model was also able to identify structural features that resulted in a lower accuracy or a higher uncertainty in the negative prediction, termed misclassified and unclassified features respectively. These features could serve as an aid for further expert assessment of the negative in silico prediction.
Related Concept Videos
Reliability and Validity
Habitat Fragmentation
Predicting Molecular Geometry
Negative Regulator Molecules
Distribution Reliability and Automation
Prediction Intervals
However, the point estimate is most likely not the exact value of the population parameter, but close to it. After calculating point estimates, we construct interval estimates, called confidence intervals or prediction intervals. This prediction interval comprises a range of values unlike the point estimate and is a better predictor of the observed sample value, y.

