Related Experiment Video
Updated: Dec 6, 2025

Watershed Planning within a Quantitative Scenario Analysis Framework
Published on: July 24, 2016
On the interpretability of predictors in spatial data science: the information horizon
Thorsten Behrens1,2, Raphael A Viscarra Rossel3
1Cluster of Excellence "Machine Learning: New Perspectives for Science", Eberhard Karls University Tuebingen, Maria von Linden Str. 6, 72076, Tübingen, Germany. thorsten.behrens@soilution.de.
Abstract:
Two important theories in spatial modelling relate to structural and spatial dependence. Structural dependence refers to environmental state-factor models, where an environmental property is modelled as a function of the states and interactions of environmental predictors, such as climate, parent material or relief. Commonly, the functions are regression or supervised classification algorithms. Spatial dependence is present in most environmental properties and forms the basis for spatial interpolation and geostatistics. In machine learning, modelling with geographic coordinates or Euclidean distance fields, which resemble linear variograms with infinite ranges, can produce similar interpolations. Interpolations do not lend themselves to causal interpretations. Conversely, with structural modeling, one can, potentially, extract knowledge from the modelling. Two important characteristics of such interpretable environmental modelling are scale and information content. Scale is relevant because very coarse scale predictors can show nearly infinite ranges, falling out of what we call the information horizon, i.e. interpretation using domain knowledge isn't possible. Regarding information content, recent studies have shown that meaningless predictors, such as paintings or photographs of faces, can be used for spatial environmental modelling of ecological and soil properties, with accurate evaluation statistics. Here, we examine under which conditions modelling with such predictors can lead to accurate statistics and whether an information horizon can be derived for scale and information content.
Related Concept Videos
Levels of Use of a GIS
Selected Data About Geographic Locations
Manipulation and Analysis
Introduction to GIS
Prediction Intervals
However, the point estimate is most likely not the exact value of the population parameter, but close to it. After calculating point estimates, we construct interval estimates, called confidence intervals or prediction intervals. This prediction interval comprises a range of values unlike the point estimate and is a better predictor of the observed sample value, y.
Outliers and Influential Points

