Related Experiment Video
Updated: Mar 5, 2026

Use of Principal Components for Scaling Up Topographic Models to Map Soil Redistribution and Soil Organic Carbon
Published on: October 16, 2018
Development of Multiple Regression Models to Predict Sources of Fecal Pollution
Abstract:
This study assessed the usefulness of multivariate statistical tools to characterize watershed dynamics and prioritize streams for remediation. Three multiple regression models were developed using water quality data collected from Sinking Creek in the Watauga River watershed in Northeast Tennessee. Model 1 included all water quality parameters, model 2 included parameters identified by stepwise regression, and model 3 was developed using canonical discriminant analysis. Models were evaluated in seven creeks to determine if they correctly classified land use and level of fecal pollution. At the watershed level, the models were statistically significant (p < 0.001) but with low r2 values (Model 1 r2 = 0.02, Model 2 r2 = 0.01, Model 3 r2 = 0.35). Model 3 correctly classified land use in five of seven creeks. These results suggest this approach can be used to set priorities and identify pollution sources, but may be limited when applied across entire watersheds.
Related Concept Videos
Multiple Regression
Farmers can use multiple regression to determine the crop yield based on more than one factor, such as water availability, fertilizer, soil properties, etc. Here, the crop yield is the response or dependent variable as it depends on the other independent variables. The analysis requires the construction of a scatter plot...
Mechanistic Models: Compartment Models in Individual and Population Analysis
Steps in Outbreak Investigation
Regression Analysis
In regression analysis, a regression equation is determined based on the line of best fit– a line that best fits the data points plotted in a graph. This line is also called the regression line. The algebraic equation for the regression line is called the regression equation. It is represented as:
Residuals and Least-Squares Property
If the observed data point lies above the line, the residual is positive, and the line underestimates the actual data value for y. If the observed data point lies below the line, the residual is negative, and the line overestimates the actual data value for y.
The process of fitting the best-fit...

