Related Experiment Video
Updated: Dec 15, 2025

Implementation of a Real-Time Psychosis Risk Detection and Alerting System Based on Electronic Health Records using CogStack
Published on: May 15, 2020
Prediction of the Number of Patients Infected with COVID-19 Based on Rolling Grey Verhulst Models
Yu-Feng Zhao1, Ming-Huan Shou1, Zheng-Xin Wang1
1School of Economics, Zhejiang University of Finance & Economics, Hangzhou 310018, China.
Abstract:
The outbreak of a novel coronavirus (SARS-CoV-2) has caused a large number of residents in China to be infected with a highly contagious pneumonia recently. Despite active control measures taken by the Chinese government, the number of infected patients is still increasing day by day. At present, the changing trend of the epidemic is attracting the attention of everyone. Based on data from 21 January to 20 February 2020, six rolling grey Verhulst models were built using 7-, 8- and 9-day data sequences to predict the daily growth trend of the number of patients confirmed with COVID-19 infection in China. The results show that these six models consistently predict the S-shaped change characteristics of the cumulative number of confirmed patients, and the daily growth decreased day by day after 4 February. The predicted results obtained by different models are very approximate, with very high prediction accuracy. In the training stage, the maximum and minimum mean absolute percentage errors (MAPEs) are 4.74% and 1.80%, respectively; in the testing stage, the maximum and minimum MAPEs are 4.72% and 1.65%, respectively. This indicates that the predicted results show high robustness. If the number of clinically diagnosed cases in Wuhan City, Hubei Province, China, where COVID-19 was first detected, is not counted from 12 February, the cumulative number of confirmed COVID-19 cases in China will reach a maximum of 60,364-61,327 during 17-22 March; otherwise, the cumulative number of confirmed cases in China will be 78,817-79,780.
Related Concept Videos
Steps in Outbreak Investigation
Statistical Methods for Analyzing Epidemiological Data
Residuals and Least-Squares Property
If the observed data point lies above the line, the residual is positive, and the line underestimates the actual data value for y. If the observed data point lies below the line, the residual is negative, and the line overestimates the actual data value for y.
The process of fitting the best-fit...
Prediction Intervals
However, the point estimate is most likely not the exact value of the population parameter, but close to it. After calculating point estimates, we construct interval estimates, called confidence intervals or prediction intervals. This prediction interval comprises a range of values unlike the point estimate and is a better predictor of the observed sample value, y.
Interpreting Run Charts
Pareto Chart
The Pareto chart is named after the Italian economist Vilfredo Pareto, who described the Pareto...

