多变量预测模型的开发和验证,以社区为基础的老年人退学计划中退学.
Masanori Morikawa1,2, Kenji Harada1, Satoshi Kurita3
1Center for Gerontology and Social Science, National Center for Geriatrics and Gerontology, Obu, Japan.
Journal of physical activity & health
|April 8, 2025
概括
一个新的模型预测,老年人会退出社区项目. 它识别了身体和认知功能等关键因素,但更可靠地预测了谁会留下,而不是谁会离开.
科学领域:
- 老年学是指老年学的学科.
- 公共卫生 公共卫生
- 预测建模预测建模
背景情况:
- 基于社区的项目对于老年人的福祉至关重要.
- 预测和预防学是计划有效性和参与者的参与度至关重要的.
研究的目的:
- 开发和验证一种多变量模型,用于预测老年人从社区外出计划中学.
- 确定这个人口群体中节目消耗的关键预测因素.
主要方法:
- 利用了来自老年综合征研究的5905名老年人的前队列.
- 在培训和验证数据集上采用极端梯度增强算法 (6:2:2比).
- 在测试数据集上使用接收器操作特征 (ROC) 和校准图表进行评估模型区分和校准.
主要成果:
- 该模型实现了ROC曲线下的面积为0.701.
- 发现的关键特征包括认知功能,身体功能,以及参与运动/体育活动的意愿.
- 该模型显示了高特异性 (0.915) 和负预测值 (0.718) 脱学预测.
结论:
- 开发的预测模型在识别不太可能放弃的参与者方面显示出可靠性.
- 身体和认知功能,以及对身体活动的意愿,似乎是计划坚持的主要预测因素.
- 可能需要进一步改进,以提高可能退出参与者的预测准确度.
相关概念视频
Prediction Intervals
2.2K
The interval estimate of any variable is known as the prediction interval. It helps decide if a point estimate is dependable.
However, the point estimate is most likely not the exact value of the population parameter, but close to it. After calculating point estimates, we construct interval estimates, called confidence intervals or prediction intervals. This prediction interval comprises a range of values unlike the point estimate and is a better predictor of the observed sample value, y.
However, the point estimate is most likely not the exact value of the population parameter, but close to it. After calculating point estimates, we construct interval estimates, called confidence intervals or prediction intervals. This prediction interval comprises a range of values unlike the point estimate and is a better predictor of the observed sample value, y.
2.2K
Multiple Regression
2.9K
Multiple regression assesses a linear relationship between one response or dependent variable and two or more independent variables. It has many practical applications.
Farmers can use multiple regression to determine the crop yield based on more than one factor, such as water availability, fertilizer, soil properties, etc. Here, the crop yield is the response or dependent variable as it depends on the other independent variables. The analysis requires the construction of a scatter plot...
Farmers can use multiple regression to determine the crop yield based on more than one factor, such as water availability, fertilizer, soil properties, etc. Here, the crop yield is the response or dependent variable as it depends on the other independent variables. The analysis requires the construction of a scatter plot...
2.9K
Comparing the Survival Analysis of Two or More Groups
92
Survival analysis is a cornerstone of medical research, used to evaluate the time until an event of interest occurs, such as death, disease recurrence, or recovery. Unlike standard statistical methods, survival analysis is particularly adept at handling censored data—instances where the event has not occurred for some participants by the end of the study or remains unobserved. To address these unique challenges, specialized techniques like the Kaplan-Meier estimator, log-rank test, and...
92
Longitudinal Research
11.8K
Sometimes we want to see how people change over time, as in studies of human development and lifespan. When we test the same group of individuals repeatedly over an extended period of time, we are conducting longitudinal research. Longitudinal research is a research design in which data-gathering is administered repeatedly over an extended period of time. For example, we may survey a group of individuals about their dietary habits at age 20, retest them a decade later at age 30, and then again...
11.8K
Regression Analysis
5.5K
Regression analysis is a statistical tool that describes a mathematical relationship between a dependent variable and one or more independent variables.
In regression analysis, a regression equation is determined based on the line of best fit– a line that best fits the data points plotted in a graph. This line is also called the regression line. The algebraic equation for the regression line is called the regression equation. It is represented as:
In regression analysis, a regression equation is determined based on the line of best fit– a line that best fits the data points plotted in a graph. This line is also called the regression line. The algebraic equation for the regression line is called the regression equation. It is represented as:
5.5K
Regression Toward the Mean
6.3K
Regression toward the mean (“RTM”) is a phenomenon in which extremely high or low values—for example, and individual’s blood pressure at a particular moment—appear closer to a group’s average upon remeasuring. Although this statistical peculiarity is the result of random error and chance, it has been problematic across various medical, scientific, financial and psychological applications. In particular, RTM, if not taken into account, can interfere when...
6.3K


