A Methodology for Validating Diversity in Synthetic Time Series Generation
Fouad Bahrpeyma1, Mark Roantree2, Paolo Cappellari3
1Insight Centre for Data Analytics, School of Computing, Dublin City University, Dublin 9, Ireland.
Abstract:
In order for researchers to deliver robust evaluations of time series models, it often requires high volumes of data to ensure the appropriate level of rigor in testing. However, for many researchers, the lack of time series presents a barrier to a deeper evaluation. While researchers have developed and used synthetic datasets, the development of this data requires a methodological approach to testing the entire dataset against a set of metrics which capture the diversity of the dataset. Unless researchers are confident that their test datasets display a broad set of time series characteristics, it may favor one type of predictive model over another. This can have the effect of undermining the evaluation of new predictive methods. In this paper, we present a new approach to generating and evaluating a high number of time series data. The construction algorithm and validation framework are described in detail, together with an analysis of the level of diversity present in the synthetic dataset.
Related Concept Videos
Empirical Method to Interpret Standard Deviation
This rule is used widely in statistics to calculate the proportion of data values...
Data Validation
Key parameters for method validation include:
Variability: Analysis
The range is a simple measure of variability, indicating the difference between the highest and...
Wald-Wolfowitz Runs Test I
The test works...
Estimating Population Standard Deviation
Sampling Methods: Overview
In analytical chemistry, the choice of...


