Related Experiment Video
Updated: Sep 18, 2025

Composition and Distribution Analysis of Bioaerosols Under Different Environmental Conditions
Published on: January 7, 2019
Source Analysis of Ozone Pollution in Liaoyuan City's Atmosphere Based on Machine Learning Models and HYSPLIT
Xinyu Zou1, Xinlong Li1, Dali Wang2
1College of New Energy and Environment, Jilin University, Changchun 130012, China.
Abstract:
Firstly, this study investigates the spatiotemporal distribution characteristics of the ozone (O3) pollution in Liaoyuan City using monitoring data from 2015 to 2024. Then, three machine learning models (ML)-random forest (RF), support vector machine (SVM), and artificial neural network (ANN)-are employed to quantify the influence of meteorological and non-meteorological factors on O3 concentrations. Finally, the HYSPLIT clustering method and CMAQ model are utilized to analyze inter-regional transport characteristics, identifying the causes of O3 pollution. The results indicate that O3 pollution in Liaoyuan exhibits a distinct seasonal pattern, with the highest concentrations found in spring and summer, peaking in the afternoon. Among the three ML models, the random forest model demonstrates the best predictive performance (R2 = 0.9043). Feature importance identifies NO2 as the primary driving factor, followed by meteorological conditions in the second quarter and land surface characteristics. Furthermore, regional transport significantly contributes to O3 pollution, with approximately 80% of air mass trajectories in heavily polluted episodes originating from adjacent industrial areas and the sea. The combined effects of transboundary precursors and O3 transport with local emissions and meteorological conditions further increase the O3 pollution level. This study highlights the need to strengthen coordinated NOX and VOCs emission reductions and enhance regional joint prevention and control strategies in China.
More Related Videos
Related Concept Videos
Cluster Sampling Method
To choose a cluster sample, divide the population into clusters (groups) and then randomly select some of the clusters. All the members from these clusters are in the cluster sample. For example, if you randomly sample four departments from your...
Sampling Plans
Random sampling is a method where each member of the population has an equal chance of being selected for the sample. It involves selecting individuals randomly, often using random number generators or lottery-type methods. For example, when analyzing the properties of a...
Precipitation and Co-precipitation
Boundary Layer Characteristics
Steps in Outbreak Investigation
Regression Analysis
In regression analysis, a regression equation is determined based on the line of best fit– a line that best fits the data points plotted in a graph. This line is also called the regression line. The algebraic equation for the regression line is called the regression equation. It is represented as:

