Applying the Kolmogorov-Zurbenko filter followed by random forest models to 7Be observations in Spain (2006-2021)
Ander Nafarrate1, Susana Petisco-Ferrero1, Raquel Idoeta1
1Energy Engineering Department, University of the Basque Country, UPV/EHU, Plaza Torres Quevedo, s/n, Bilbao, 48013, Spain.
Abstract:
In this study, we analysed 7Be weekly surface measurements from six Spanish laboratories from 2006 to 2021. The Kolmogorov-Zurbenko filter was applied to the six 7Be time series, and following an iterative process, the original data were divided into two fractions: one related to variations characterized by periods above 33 days (including, among others, the seasonal cycle) and the second noisier fraction related to mechanisms originating from variations with periods below 33 days. Both fractions were independent at the six locations. The second machine-based step using random forest models was applied with the aim of identifying the most influential inputs to the observed 7Be concentrations, and machine learning-inspired regression models were fitted. With respect to seasonal components, the results indicated that the memory of the system was the most influential input, as expected by the large fraction of variance explained by the seasonal cycle, followed by that of humidity and wind-related variables. For the fraction corresponding to periods below 33 d, precipitation-, humidity-, and radiation-related variables were the most influential. This methodology has made it possible to successfully describe the major mechanisms known to be involved in the generation of the surface 7Be concentrations observed in Spain.
Related Concept Videos
Statistical Methods for Analyzing Epidemiological Data
Classification of Signals
A continuous-time signal holds a value at every instant in time, representing information seamlessly. In contrast, a discrete-time signal holds values only at specific moments, often denoted as x(n), where...
Random Variables
Uppercase letters such as X or Y denote a random variable. Lowercase letters like x or y denote the value of a random variable. If X is a random variable, then X is written in words, and x is given as a number.
For example, let X = the...
Prediction Intervals
However, the point estimate is most likely not the exact value of the population parameter, but close to it. After calculating point estimates, we construct interval estimates, called confidence intervals or prediction intervals. This prediction interval comprises a range of values unlike the point estimate and is a better predictor of the observed sample value, y.
Steps in Outbreak Investigation
Econometric Views (EViews)


