Related Experiment Video
Updated: Mar 24, 2026

Author Spotlight: UAV Remote Sensing for Efficient Invasive Plant Biomass Estimation
Published on: February 9, 2024
Satellite-Based NO2 and Model Validation in a National Prediction Model Based on Universal Kriging and Land-Use
Michael T Young1, Matthew J Bechle2, Paul D Sampson3
1Department of Epidemiology, University of Washington 4225 Roosevelt Way NE, Seattle, Washington 98105, United States.
Abstract:
Epidemiological studies increasingly rely on exposure prediction models. Predictive performance of satellite data has not been evaluated in a combined land-use regression/spatial smoothing context. We performed regionalized national land-use regression with and without universal kriging on annual average NO2 measurements (1990-2012, contiguous U.S. EPA sites). Regression covariates were dimension-reduced components of 418 geographic variables including distance to roadway. We estimated model performance with two cross-validation approaches: using randomly selected groups and, in order to assess predictions to unmonitored areas, spatially clustered cross-validation groups. Ground-level NO2 was estimated from satellite-derived NO2 and was assessed as an additional regression covariate. Kriging models performed consistently better than nonkriging models. Among kriging models, conventional cross-validated R(2) (R(2)cv) averaged over all years was 0.85 for the satellite data models and 0.84 for the models without satellite data. Average spatially clustered R(2)cv was 0.74 for the satellite data models and 0.64 for the models without satellite data. The addition of either kriging or satellite data to a well-specified NO2 land-use regression model each improves prediction. Adding the satellite variable to a kriging model only marginally improves predictions in well-sampled areas (conventional cross-validation) but substantially improves predictions for points far from monitoring locations (clustered cross-validation).
More Related Videos
12:26Integrating Remote Sensing with Species Distribution Models; Mapping Tamarisk Invasions Using the Software for Assisted Habitat Modeling SAHM
Published on: October 11, 2016
09:44Use of Principal Components for Scaling Up Topographic Models to Map Soil Redistribution and Soil Organic Carbon
Published on: October 16, 2018
Related Concept Videos
Regression Analysis
In regression analysis, a regression equation is determined based on the line of best fit– a line that best fits the data points plotted in a graph. This line is also called the regression line. The algebraic equation for the regression line is called the regression equation. It is represented as:
Multiple Regression
Farmers can use multiple regression to determine the crop yield based on more than one factor, such as water availability, fertilizer, soil properties, etc. Here, the crop yield is the response or dependent variable as it depends on the other independent variables. The analysis requires the construction of a scatter plot...
Residuals and Least-Squares Property
If the observed data point lies above the line, the residual is positive, and the line underestimates the actual data value for y. If the observed data point lies below the line, the residual is negative, and the line overestimates the actual data value for y.
The process of fitting the best-fit...
Levels of Use of a GIS
Prediction Intervals
However, the point estimate is most likely not the exact value of the population parameter, but close to it. After calculating point estimates, we construct interval estimates, called confidence intervals or prediction intervals. This prediction interval comprises a range of values unlike the point estimate and is a better predictor of the observed sample value, y.