对于体育医学和运动科学中的样本大小计算的ChatGPT:一个警告说明
Jabeur Methnani1,2, Imed Latiri3, Ismail Dergaa4,5,6
1LR19ES09, Laboratoire de Physiologie de l'Exercice et Physiopathologie: de l'Intégré au Moléculaire "Biologie, Médecine et Santé," Faculty of Medicine of Sousse, University of Sousse, Sousse,Tunisia.
在四项体育科学研究中,ChatGPT只准确计算了其中一项样本大小,突出了人工智能驱动的研究计算中的潜在错误. 这些新兴工具需要进一步验证.
科学领域:
- 运动科学 运动科学
- 运动医学 运动医学
- 生物统计学 生物统计学
背景情况:
- 准确的样本大小计算对于体育科学和体育医学研究研究的有效性和统计能力至关重要.
- 像ChatGPT这样的大型语言模型 (LLM) 提供了帮助研究人员的潜力,但它们对于复杂的统计任务的可靠性需要进行调查.
研究的目的:
- 评估ChatGPT在计算体育科学和体育医学研究的样本大小时的准确性.
- 在LLM生成的样本大小计算中识别潜在的错误和不一致.
主要方法:
- 分析了四篇发表的体育科学/医学研究论文,研究设计不同.
- 将所有必要的统计数据 (平均值,SD值,Z值) 和研究设计细节输入ChatGPT.
- 重新提示ChatGPT提供相同的信息以评估响应可重现性.
主要成果:
- 在四项研究中,ChatGPT仅对其中一项 (随机对照试验) 正确计算了样本大小.
- 该LLM未能正确确定调查论文样本大小计算的适当公式.
- 对一个示例重复使用相同的提示结果产生了不同的样本大小计算,表明不一致.
结论:
- 像ChatGPT这样的当前LLM在样本大小计算中可能会产生错误和不一致,即使输入完整且准确.
- 研究人员在使用人工智能工具进行统计计算和验证结果时必须谨慎使用.
- 未来的研究应该探索更先进的AI模型及其支持科学研究任务的能力.
更多相关视频
05:51Assessing the Accuracy of Fitness Smartwatch Data for Cardiovascular and Physical Activity Monitoring: A Validation Study in Digital Health
Published on: February 21, 2025
06:13Author Spotlight: Exploring the Impact of Reduced Resistance Exercise Volume on Metabolic Health
Published on: December 1, 2023
相关概念视频
Sample Size Calculation
The sample size for the given experiment or sampling effort is fundamental to any study design. Sample size decides the number of...
One-Way ANOVA: Unequal Sample Sizes
Contaminants and Errors
Another key consideration is determining the appropriate number of samples required to...
One-Way ANOVA: Equal Sample Sizes
Different sample means can result in different values for the variance estimate: variance between samples. This is because the variance between samples is calculated as the product of the sample size and the variance between the...
Bootstrapping
Estimating Population Mean with Unknown Standard Deviation
William S. Gosset (1876–1937) of the...
