Spatial interpolation of regional PM2.5 concentrations in China during COVID-19 incorporating multivariate data
Pengzhi Wei1,2, Shaofeng Xie3,4, Liangke Huang3,4
1Institute of Artificial Intelligence, School of Computer Science, Wuhan University, Wuhan, 430072, China.
Abstract:
During specific periods when the PM2.5 variation pattern is unusual, such as during the coronavirus disease 2019 (COVID-19) outbreak, epidemic PM2.5 regional interpolation models have been relatively little investigated, and little consideration has been given to the residuals of optimized models and changes in model interpolation accuracy for the PM2.5 concentration under the influence of epidemic phenomena. Therefore, this paper mainly introduces four interpolation methods (kriging, empirical Bayesian kriging, tensor spline function and complete regular spline function), constructs geographically weighted regression (GWR) models of the PM2.5 concentration in Chinese regions for the periods from January-June 2019 and January-June 2020 by considering multiple factors, and optimizes the GWR regression residuals using these four interpolation methods, thus achieving the purpose of enhancing the model accuracy. The PM2.5 concentrations in many regions of China showed a downward trend during the same period before and after the COVID-19 outbreak. Atmospheric pollutants, meteorological factors, elevation, zenith wet delay (ZWD), normalized difference vegetation index (NDVI) and population maintained a certain relationship with the PM2.5 concentration in terms of linear spatial relationships, which could explain why the PM2.5 concentration changed to a certain extent. By evaluating the model accuracy from two perspectives, i.e., the overall interpolation effect and the validation set interpolation effect, the results showed that all four interpolation methods could improve the numerical accuracy of GWR to different degrees, among which the tensor spline function and the fully regular spline function achieved the most stable effect on the correction of GWR residuals, followed by kriging and empirical Bayesian kriging.
Related Concept Videos
Residuals and Least-Squares Property
If the observed data point lies above the line, the residual is positive, and the line underestimates the actual data value for y. If the observed data point lies below the line, the residual is negative, and the line overestimates the actual data value for y.
The process of fitting the best-fit...
Statistical Methods for Analyzing Epidemiological Data
Single Nucleotide Polymorphisms-SNPs
Sampling Plans
Random sampling is a method where each member of the population has an equal chance of being selected for the sample. It involves selecting individuals randomly, often using random number generators or lottery-type methods. For example, when analyzing the properties of a...


