曲线的情况:用预测器的二次和三次多项式函数进行参数回归应该是常规的
Edward Kroc1, Oscar L Olvera Astivia2
1Department of Educational and Counselling Psychology, University of British Columbia.
Psychological methods
|December 14, 2023
概括
多项式回归,特别是与二次和三次项,被推作为应用从业者和学生的默认建模技术. 它比复杂的非参数方法提供了更高的解释性和灵活性,用于建模非线性.
科学领域:
- 统计 统计 统计 统计
- 数据建模数据建模
- 回归分析是一种回归分析.
背景情况:
- 多项式回归是一种广泛讨论的,但建议的统计建模技术.
- 非参数替代方案通常需要复杂的数学理解,这给非统计学家带来了挑战.
研究的目的:
- 倡导将低阶多项式回归 (二级和三级项) 纳入标准模型构建工具箱.
- 将多项式回归定位为向学生和从业人员教授非线性建模的默认方法.
- 为非统计观众强调多项式回归比非参数方法的优势.
主要方法:
- 该研究基于其可解释性,灵活性和易用性,为多项式回归提出了理由.
- 将多项式回归与非参数式替代方案进行比较,强调其对应用实践者的实际优势.
- 讨论低阶多项式回归对特定效应 (如地板/天花板效应和局部线性) 的建模能力.
主要成果:
- 对于非统计学家来说,使用二次和三次项的多项式回归被认为优于非参数方法.
- 它有效地模拟非线性,全球和本地预测效应,地板/天花板效应和本地线性.
- 该方法可以防止预测者之间的虚假相互作用效应的推断.
结论:
- 低阶多项式回归应该是模拟非线性性的默认技术,因为它具有实际优势.
- 现有的反对多项式回归的论点通常基于误解或过时的信息.
- 这种技术使非统计学家能够构建现实的和可解释的模型.
相关概念视频
Multiple Regression
3.0K
Multiple regression assesses a linear relationship between one response or dependent variable and two or more independent variables. It has many practical applications.
Farmers can use multiple regression to determine the crop yield based on more than one factor, such as water availability, fertilizer, soil properties, etc. Here, the crop yield is the response or dependent variable as it depends on the other independent variables. The analysis requires the construction of a scatter plot...
Farmers can use multiple regression to determine the crop yield based on more than one factor, such as water availability, fertilizer, soil properties, etc. Here, the crop yield is the response or dependent variable as it depends on the other independent variables. The analysis requires the construction of a scatter plot...
3.0K
Parametric Survival Analysis: Weibull and Exponential Methods
445
Parametric survival analysis models survival data by assuming a specific probability distribution for the time until an event occurs. The Weibull and exponential distributions are two of the most commonly used methods in this context, due to their versatility and relatively straightforward application.
Weibull Distribution
The Weibull distribution is a flexible model used in parametric survival analysis. It can handle both increasing and decreasing hazard rates, depending on its shape parameter...
Weibull Distribution
The Weibull distribution is a flexible model used in parametric survival analysis. It can handle both increasing and decreasing hazard rates, depending on its shape parameter...
445
Residuals and Least-Squares Property
7.4K
The vertical distance between the actual value of y and the estimated value of y. In other words, it measures the vertical distance between the actual data point and the predicted point on the line
If the observed data point lies above the line, the residual is positive, and the line underestimates the actual data value for y. If the observed data point lies below the line, the residual is negative, and the line overestimates the actual data value for y.
The process of fitting the best-fit...
If the observed data point lies above the line, the residual is positive, and the line underestimates the actual data value for y. If the observed data point lies below the line, the residual is negative, and the line overestimates the actual data value for y.
The process of fitting the best-fit...
7.4K
Regression Analysis
5.7K
Regression analysis is a statistical tool that describes a mathematical relationship between a dependent variable and one or more independent variables.
In regression analysis, a regression equation is determined based on the line of best fit– a line that best fits the data points plotted in a graph. This line is also called the regression line. The algebraic equation for the regression line is called the regression equation. It is represented as:
In regression analysis, a regression equation is determined based on the line of best fit– a line that best fits the data points plotted in a graph. This line is also called the regression line. The algebraic equation for the regression line is called the regression equation. It is represented as:
5.7K
Variation
6.8K
An important characteristic of any set of data is the variation in the data. In some data sets, the data values are concentrated closely near the mean; in other data sets, the data values are more widely spread out from the mean. The most common measure of variation, or spread, is the standard deviation, which is the square root of variance.
When independent and dependent variables are plotted on a scatter plot, the slope of a line is a value that describes the rate of change between the two...
When independent and dependent variables are plotted on a scatter plot, the slope of a line is a value that describes the rate of change between the two...
6.8K
Mechanistic Models: Compartment Models in Algorithms for Numerical Problem Solving
56
Mechanistic models play a crucial role in algorithms for numerical problem-solving, particularly in nonlinear mixed effects modeling (NMEM). These models aim to minimize specific objective functions by evaluating various parameter estimates, leading to the development of systematic algorithms. In some cases, linearization techniques approximate the model using linear equations.
In individual population analyses, different algorithms are employed, such as Cauchy's method, which uses a...
In individual population analyses, different algorithms are employed, such as Cauchy's method, which uses a...
56


