Related Experiment Video
Updated: Jan 28, 2026

Comparison of Three Different Methods for Determining Cell Proliferation in Breast Cancer Cell Lines
Published on: September 3, 2016
A comparison of statistical learning methods for deriving determining factors of accident occurrence from an
Matthias Schlögl1, Rainer Stütz2, Gregor Laaha3
1Institute of Applied Statistics and Scientific Computing, University of Natural Resources and Life Sciences, Vienna, Austria; Transportation Infrastructure Technologies, Austrian Institute of Technology, Vienna, Austria.
Abstract:
One of the main aims of accident data analysis is to derive the determining factors associated with road traffic accident occurrence. While current studies mainly use variants of count data regression to achieve this aim, the problem can also be considered as a binary classification task, with the dichotomous target variable indicating events (accidents) and non-events (no accidents). The effects of 45 variables - describing road condition and geometry, traffic volume and regulations, weather, and accident time - are analyzed using a dataset in high temporal (1 h) and spatial (250 m) resolution, covering the whole highway network of Austria over the period of four consecutive years. A combination of synthetic minority oversampling and maximum dissimilarity undersampling is used to balance the training dataset. We employ and compare a series of statistical learning techniques with respect to their predictive performance and discuss the importance of determining factors of accident occurrence from the ensemble of models. Findings substantiate that a trade-off between accuracy and sensitivity is inherent to imbalanced classification problems. Results show satisfying performance of tree-based methods which exhibit accuracies between 75% and 90% while exhibiting sensitivities between 30% and 50%. Overall, this analysis emphasizes the merits of using high-resolution data in the context of accident analysis.
Related Concept Videos
Statistical Significance
Statistical Methods for Analyzing Epidemiological Data
Statistical Methods to Analyze Parametric Data: ANOVA
One-way ANOVA is applied when a single independent variable or factor is scrutinized. It compares...
The Sense of Self: Reflected Self-Appraisal and Social Comparison
Probability in Statistics
An example of a simple event is a coin toss. The result of a coin toss is either a head or a tail. Here, head and tail are two simple events. These two simple events make up the sample space. Further, the probability of an event occurring falls within the range of 0 to 1. The probability of an...
Introduction to Statistics
In statistics, the collection of individuals or objects under study is called population. The idea of sampling is to select a portion of the larger population...

