Related Experiment Video
Updated: Nov 10, 2025

Author Spotlight: Advancements in Multiplex Detection of Respiratory Viruses
Published on: November 10, 2023
Forecasting COVID-19 Confirmed Cases Using Empirical Data Analysis in Korea
Da Hye Lee1, Youn Su Kim1, Young Youp Koh2
1Department of Computer Science and Statistics, Chosun University, Gwangju, 61452, Korea.
Abstract:
From November to December 2020, the third wave of COVID-19 cases in Korea is ongoing. The government increased Seoul's social distancing to the 2.5 level, and the number of confirmed cases is increasing daily. Due to a shortage of hospital beds, treatment is difficult. Furthermore, gatherings at the end of the year and the beginning of next year are expected to worsen the effects. The purpose of this paper is to emphasize the importance of prediction timing rather than prediction of the number of confirmed cases. Thus, in this study, five groups were set according to minimum, maximum, and high variability. Through empirical data analysis, the groups were subdivided into a total of 19 cases. The cumulative number of COVID-19 confirmed cases is predicted using the auto regressive integrated moving average (ARIMA) model and compared with the actual number of confirmed cases. Through group and case-by-case prediction, forecasts can accurately determine decreasing and increasing trends. To prevent further spread of COVID-19, urgent and strong government restrictions are needed. This study will help the government and the Korea Disease Control and Prevention Agency (KDCA) to respond systematically to a future surge in confirmed cases.
Related Concept Videos
Steps in Outbreak Investigation
Statistical Methods for Analyzing Epidemiological Data
Pie Chart
In a pie chart, the central angle, the arc length of each slice, and the area are directly proportional to the quantity or percentage it represents. Some real-world examples that can be depicted using pie charts include marks obtained by students...
Residuals and Least-Squares Property
If the observed data point lies above the line, the residual is positive, and the line underestimates the actual data value for y. If the observed data point lies below the line, the residual is negative, and the line overestimates the actual data value for y.
The process of fitting the best-fit...
Pareto Chart
The Pareto chart is named after the Italian economist Vilfredo Pareto, who described the Pareto...
Introduction to Epidemiology

