Assessing the accuracy of various statistical models for forecasting PM : a case study from diverse regions of
Sajeed I Ghanchi1, Dishant M Pandya2, Manan Shah3
1Department of Mathematics, Pandit Deendayal Energy University, Gandhinagar, 382426, Gujarat, India.
Abstract:
PM is the most hazardous air pollutant due to its smaller size, which allows deeper bodily penetration. Three diverse regions from Gujarat, India, namely Sector 10, Maninagar, and Vatva, which have green space, high population concentration, and industries, respectively, were chosen to forecast PM concentration for the next day. Four statistical models, including Multiple Linear Regression (MLR), Principal Component Regression (PCR), Simple Exponential Smoothing (SES), and Autoregressive Integrated Moving Average (ARIMA), were chosen to forecast PM levels. For this study, data of various pollutants and meteorological parameters were collected from February 2019 to September 2023. Analysis of the seasonal patterns of PM revealed elevated concentrations during post-monsoon and winter, in contrast to reduced levels during summer and monsoon. Statistical analysis revealed that the concentration of PM in Sector 10 is much lower than in the other two regions. The analysis of the test results, utilising various accuracy measures like RMSE, MAE, MAPE, IA, and others, indicated that Sector 10 achieved the highest precision in its results. While assessing the models' accuracy on the test data, the ARIMA model demonstrated the highest level of precision. The average RMSE, MAE, and MAPE values for the ARIMA model were 12.63, 8.59, and 0.24, respectively. In the comparison of the performance between these statistical models and the neural network-based Multilayer Perceptron (MLP) model, it was observed that the statistical model demonstrated superior performance over the MLP model.
More Related Videos
12:26Integrating Remote Sensing with Species Distribution Models; Mapping Tamarisk Invasions Using the Software for Assisted Habitat Modeling SAHM
Published on: October 11, 2016
09:03Assessment of Methane and Nitrous Oxide Fluxes from Paddy Field by Means of Static Closed Chambers Maintaining Plants Within Headspace
Published on: September 6, 2018
Related Concept Videos
Random Error
Testing a Claim about Standard Deviation
The hypothesis testing for the claim of population standard deviation (or variance) requires the data and samples to be random and unbiased. The population distribution also must be normal. There is no specific requirement on the sample size as the estimation is based on the chi-square distribution.
As a first step, the hypothesis (null and alternative) concerning the claim about...
Estimating Population Standard Deviation
Estimating Population Mean with Unknown Standard Deviation
William S. Gosset (1876–1937) of the...
Estimating Population Mean with Known Standard Deviation
The confidence interval estimate will have the form as follows:
(point estimate - error bound, point estimate +...
Statistical Methods for Analyzing Epidemiological Data
