Related Concept Videos
Residuals and Least-Squares Property
If the observed data point lies above the line, the residual is positive, and the line underestimates the actual data value for y. If the observed data point lies below the line, the residual is negative, and the line overestimates the actual data value for y.
The process of fitting the best-fit...
Quantifying and Rejecting Outliers: The Grubbs Test
Parametric Survival Analysis: Weibull and Exponential Methods
Weibull Distribution
The Weibull distribution is a flexible model used in parametric survival analysis. It can handle both increasing and decreasing hazard rates, depending on its shape parameter...
Prediction Intervals
However, the point estimate is most likely not the exact value of the population parameter, but close to it. After calculating point estimates, we construct interval estimates, called confidence intervals or prediction intervals. This prediction interval comprises a range of values unlike the point estimate and is a better predictor of the observed sample value, y.
Quantitative Analysis
In quantitative analysis, two key measurements are made: the sample quantity and a property proportional to the amount of the analyte (the substance being analyzed). This forms the basis of the...
Statistical Methods for Analyzing Epidemiological Data
You might also read
Related Articles
Articles linked to this work by shared authors, journal, and citation graph.
Direct assessment of health impacts on hospital admission from traffic intensity in Madrid.
Related Experiment Video
Updated: Nov 3, 2025

Establishing a Competing Risk Regression Nomogram Model for Survival Data
Published on: October 23, 2020
Comparing quantile regression methods for probabilistic forecasting of NO2 pollution levels.
Sebastien Pérez Vasseur1, José L Aznarte2
1Artificial Intelligence Department, Universidad Nacional de Educación a Distancia - UNED, c/Juan del Rosal, 16, Madrid, Spain.
Forecasting air quality is crucial for managing traffic restrictions. This study compared probabilistic models for nitrogen dioxide (NO2) prediction, finding quantile gradient boosted trees performed best, though simpler models offered comparable results with less complexity.
More Related Videos
Area of Science:
- Environmental Science
- Atmospheric Chemistry
- Data Science
Background:
- Authorities use traffic restrictions to manage high nitrogen dioxide (NO2) concentrations.
- Accurate forecasting of NO2 levels is essential for timely intervention.
- Probabilistic forecasting offers advantages over point-forecasting for predicting pollutant distributions.
Purpose of the Study:
- To compare the performance of 10 state-of-the-art quantile regression models for NO2 concentration forecasting.
- To evaluate these models for predicting the full distribution of NO2 concentrations.
- To identify optimal probabilistic models for urban air quality management.
Main Methods:
- Utilized 10 quantile regression models to predict NO2 concentration distributions.
- Applied a semi-parametric approach by deriving distribution parameters from predicted quantiles.
- Evaluated models for forecasting horizons up to 60 hours in an urban setting.
Main Results:
- Quantile gradient boosted trees demonstrated superior performance in predicting both point values and full NO2 distributions.
- Quantile k-nearest neighbors with linear regression achieved comparable results with significantly reduced training time and complexity.
- The study provides a comprehensive comparison of probabilistic models for NO2 forecasting.
Conclusions:
- Probabilistic forecasting models are effective for predicting NO2 concentration exceedances and pollution peaks.
- Quantile gradient boosted trees are highly effective, but simpler models like quantile k-nearest neighbors offer a practical alternative.
- The findings support the use of advanced statistical methods for proactive air quality management.

