Related Experiment Video
Updated: Jul 13, 2025

A Method of Trigonometric Modelling of Seasonal Variation Demonstrated with Multiple Sclerosis Relapse Data
Published on: December 9, 2015
An interpretable time series machine learning method for varying forecast and nowcast lengths in wastewater-based
Mallory Lai1, Shaun S Wulff1, Yongtao Cao2
1Department of Mathematics and Statistics, University of Wyoming, 1000 E University Ave, Laramie, WY, USA.
Abstract:
Wastewater-based epidemiology has emerged as a viable tool for monitoring disease prevalence in a population. This paper details a time series machine learning (TSML) method for predicting COVID-19 cases from wastewater and environmental variables. The TSML method utilizes a number of techniques to create an interpretable, hypothesis-driven framework for machine learning that can handle different nowcast and forecast lengths. Some of the techniques employed include:•Feature engineering to construct interpretable features, like site-specific lead times, hypothesized to be potential predictors of COVID-19 cases.•Feature selection to identify features with the best predictive performance for the tasks of nowcasting and forecasting.•Prequential evaluation to prevent data leakage while evaluating the performance of the machine learning algorithm.
Related Concept Videos
Steps in Outbreak Investigation
Noncompartmental Analysis: Mean Residence Time
After the administration of a drug through intravenous bolus injection, the drug molecules are distributed throughout the body and remain there for varying periods. The MRT represents the average time these drug molecules stay in the...
Statistical Methods for Analyzing Epidemiological Data
Mechanistic Models: Compartment Models in Individual and Population Analysis
Introduction To Survival Analysis
The primary goal of survival analysis is to estimate survival time—the time...
Prediction Intervals
However, the point estimate is most likely not the exact value of the population parameter, but close to it. After calculating point estimates, we construct interval estimates, called confidence intervals or prediction intervals. This prediction interval comprises a range of values unlike the point estimate and is a better predictor of the observed sample value, y.

