Related Experiment Video
Updated: Nov 30, 2025

Composition and Distribution Analysis of Bioaerosols Under Different Environmental Conditions
Published on: January 7, 2019
Predicting PM2.5 in Well-Mixed Indoor Air for a Large Office Building Using Regression and Artificial Neural Network
Brent Lagesse1, Shuoqi Wang2, Timothy V Larson2
1Division of Computing and Software Systems, University of Washington Bothell, Bothell, Washington 98011, United States.
Abstract:
Although the exposure to PM2.5 has serious health implications, indoor PM2.5 monitoring is not a widely applied practice. Regulations on the indoor PM2.5 level and measurement schemes are not well established. Compared to other indoor settings, PM2.5 prediction models for large office buildings are particularly lacking. In response to these challenges, statistical models were developed in this paper to predict the PM2.5 concentration in well-mixed indoor air in a commercial office building. The performances of different modeling methods, including multiple linear regression (MLR), partial least squares regression (PLS), distributed lag model (DLM), least absolute shrinkage selector operator (LASSO), simple artificial neural networks (ANN), and long-short term memory (LSTM), were compared. Various combinations of environmental and meteorological parameters were used as predictors. The root-mean-square error (RMSE) of the predicted hourly PM2.5 was 1.73 μg/m3 for the LSTM model and in the range of 2.20-4.71 μg/m3 for the other models when regulatory ambient PM2.5 data were used as predictors. The LSTM models outperformed other modeling approaches across the performance metrics used by learning the predictors' temporal patterns. Even without any ambient PM2.5 information, the developed models still demonstrated relatively high skill in predicting the PM2.5 levels in well-mixed indoor air.
More Related Videos
09:33Visualizing Field Data Collection Procedures of Exposure and Biomarker Assessments for the Household Air Pollution Intervention Network Trial in India
Published on: December 23, 2022
04:04Asthma Detection Research Based on Voice Signal Processing and Machine Learning
Published on: July 22, 2025
Related Concept Videos
Regression Analysis
In regression analysis, a regression equation is determined based on the line of best fit– a line that best fits the data points plotted in a graph. This line is also called the regression line. The algebraic equation for the regression line is called the regression equation. It is represented as:
Measurement of Air Content in Concrete
The pressure method,...
Residuals and Least-Squares Property
If the observed data point lies above the line, the residual is positive, and the line underestimates the actual data value for y. If the observed data point lies below the line, the residual is negative, and the line overestimates the actual data value for y.
The process of fitting the best-fit...
Multiple Regression
Farmers can use multiple regression to determine the crop yield based on more than one factor, such as water availability, fertilizer, soil properties, etc. Here, the crop yield is the response or dependent variable as it depends on the other independent variables. The analysis requires the construction of a scatter plot...