Interpreting Neural Network Models for Toxicity Prediction by Extracting Learned Chemical Features

Moritz Walter1, Samuel J Webb2, Valerie J Gillet1

  • 1Information School, University of Sheffield, The Wave, 2 Whitham Road, Sheffield S10 2AH, U.K.

Summary

This study introduces a new method for interpreting neural network toxicity predictions by identifying chemical substructures that activate hidden neurons. This approach enhances understanding of complex models and complements existing feature attribution techniques.