Related Experiment Video
Updated: Jun 5, 2025

Measuring Attention and Visual Processing Speed by Model-based Analysis of Temporal-order Judgments
Published on: January 23, 2017
[Prediction of PM2.5 Concentration by Transformer Model Based on Attention Mechanism]
Min-Yi Liu1, Bo-Wen Cui1, Yu-Kun Wang1
1School of Ecological Environment and Urban Construction, Fujian University of Technology, Fuzhou 350118, China.
Abstract:
An improved transformer model based on a multi-head attention mechanism was constructed for long-term prediction of PM2.5 concentration. The monitoring data and meteorological data from 12 air monitoring stations in Beijing from March 2013 to December 2016 were collected and used in the transformer model. The Pearson coefficient was used to explore the key factors affecting PM2.5 concentration. A convolutional neural network model (ResNet50) and long short-term memory network model (LSTM) were introduced for comparison and explanatory variance (EVS), coefficient of determination (R2), mean square error (MSE), and mean absolute error (MAE) were selected to evaluate the performance of the model. The Pearson coefficient results showed that PM10, SO2, and NO2, CO, and atmospheric pressure (PRES) were highly correlated with PM2.5 concentration, and dew point temperature (DEWP) was strongly correlated with PM2.5 concentration, which was consistent with the preference setting of the model's automatic screening. The MSE and R2 of the transformer model are 0.009 μg·m-3 and 0.925, respectively, which decreased by 91.09% and 30.77% of MSE and increased by 38.05% and 4.65% of R2, respectively, when compared with ResNet50 and LSTM. The transformer model could capture short-term pollution changes caused by sudden changes in meteorological conditions and long-term trends with significant seasonal changes. The fitting effect of the transformer was excellent among several models, providing a novel method for long-term prediction of PM2.5 concentration. In addition, ablation experiments revealed that the increase in R2 of the transformer was relatively small after data input or manually setting preferences, with only a 2.31% and 1.51% increase, respectively, indicating that the transformer model had strong anti-interference ability for PM10 homologous data.
Related Concept Videos
Transformers with Off-Nominal Turns Ratios
Equivalent Circuits for Practical Transformers
In a practical transformer, each winding exhibits resistance and leakage reactance. The...
Three-Winding Transformers
In the per-unit equivalent circuit of a grounded Y-Y three-phase...
Transformers
The iron core has a substantial relative permeability. Therefore, the magnetic field lines generated due to the current in one winding are almost entirely confined within the core, such that the same magnetic flux permeates each turn of both...
Energy Losses in Transformers
There are four main reasons for energy losses in transformers.
The first cause can be the high resistance of the...
Reducing Line Loss
With a step-up transformer at the source, the voltage is increased, thereby reducing the current in the transmission lines since power loss...

