质量,正确性和相似性的比较 聊天GPT生成和人类撰写的基础研究摘要:跨部分研究
Shu-Li Cheng1, Shih-Jen Tsai2,3, Ya-Mei Bai2,3
1Department of Nursing, Mackay Medical College, Taipei, Taiwan.
Journal of medical Internet research
|December 25, 2023
概括
像ChatGPT这样的人工智能 (AI) 模型可以帮助研究人员,但基本研究论文的摘要质量和准确性低于人类撰写的论文. 专家可以高准确地识别人工智能生成的内容.
科学领域:
- 生物医学研究的研究.
- 科学传播是科学传播.
- 人工智能应用的人工智能应用.
背景情况:
- 聊天GPT显示了作为一个研究助理组织思想和总结发现的潜力.
- 有限的研究评估了来自全文研究论文的ChatGPT生成摘要的质量,相似性和准确性.
研究的目的:
- 评估人工智能 (AI) 模型在生成基本临床前研究论文摘要方面的有效性.
主要方法:
- 从著名的科学期刊中选择了30篇基础研究论文.
- 全文 (不包括摘要) 被输入到ChatPDF (基于ChatGPT的应用程序) 中以生成摘要.
- 八位专家盲目评估了抽象质量 (利克尔特尺度0-10),与原始摘要的相似性和准确性,还确定了人工智能生成的内容.
主要成果:
- 与原始摘要 (平均8.09) 相比,ChatGPT生成的摘要质量明显较低 (平均4.72).
- 在非结构化摘要中,质量差异比结构化格式更明显.
- 3个摘要包含错误的结论,10个被认为是人工智能生成的;评论员正确地识别了93%的准确性.
结论:
- 虽然ChatGPT生成的摘要可能不会显著改变与原始人类文本的相似性,但它们的质量是不理想的.
- 人工智能生成的摘要的准确性并不总是百分之百,因此需要对人类进行仔细的审查.
关键词:
人工智能生成的科学内容聊天GPT 聊天GPT 聊天在法学士 (LLM) 课程中.在NLP中,我们使用了NLP.一个抽象的抽象的抽象.摘要 摘要 摘要 摘要 摘要学术研究学术研究人工智能的人工智能是人工智能.提取物 提取物提取物提取 提取 提取 提取一代又一代,一代又一代的世代.产生性的产生性.语言模型语言模型语言模型语言模型自然语言处理自然语言处理.这是抄袭,抄袭.出版 出版 出版 出版 出版出版物 出版物出版物科学研究科学研究文本 文本 文本文本性 文本性 文本性更多相关视频
09:35A Protocol for Using Gene Set Enrichment Analysis to Identify the Appropriate Animal Model for Translational Research
Published on: August 16, 2017
17.9K
00:08A Cross-Disciplinary and Multi-Modal Experimental Design for Studying Near-Real-Time Authentic Examination Experiences
Published on: September 4, 2019
7.0K
相关概念视频
Comparing the Survival Analysis of Two or More Groups
195
Survival analysis is a cornerstone of medical research, used to evaluate the time until an event of interest occurs, such as death, disease recurrence, or recovery. Unlike standard statistical methods, survival analysis is particularly adept at handling censored data—instances where the event has not occurred for some participants by the end of the study or remains unobserved. To address these unique challenges, specialized techniques like the Kaplan-Meier estimator, log-rank test, and...
195
Group Design
8.9K
The most basic experimental design involves two groups: the experimental group and the control group. The two groups are designed to be the same except for one difference— experimental manipulation. The experimental group gets the experimental manipulation—that is, the treatment or variable being tested—and the control group does not. Since experimental manipulation is the only difference between the experimental and control groups, we can be sure that any differences between...
8.9K
Cross-Sectional Research
11.3K
In cross-sectional research, a researcher compares multiple segments of the population at the same time. If they were interested in people's dietary habits, the researcher might directly compare different groups of people by age. Instead of following a group of people for 20 years to see how their dietary habits changed from decade to decade, the researcher would study a group of 20-year-old individuals and compare them to a group of 30-year-old individuals and a group of 40-year-old...
11.3K
Types of Biopharmaceutical Studies: Controlled and Non-Controlled Approaches
130
Biopharmaceutical studies constitute a vital field aiming to enhance drug delivery methods and refine therapeutic approaches, drawing upon diverse interdisciplinary knowledge. In research methodologies, the choice between controlled and non-controlled studies significantly influences the study's reliability and accuracy.
Non-controlled studies, commonly employed for initial exploration, lack a control group, rendering them susceptible to biases and external influences. In contrast,...
Non-controlled studies, commonly employed for initial exploration, lack a control group, rendering them susceptible to biases and external influences. In contrast,...
130
Multiple Comparison Tests
3.9K
Multiple comparison test, abbreviated as MCT, is a post hoc analysis generally performed after comparing multiple samples with one or more tests. An MCT will help identify a significantly different sample among multiple samples or a factor among multiple factors.
It would be easy to compare two samples using a significance alpha level of 0.05. In other words, there is only one sample pair to be compared. However, it would be difficult to identify a significantly different sample if the number...
It would be easy to compare two samples using a significance alpha level of 0.05. In other words, there is only one sample pair to be compared. However, it would be difficult to identify a significantly different sample if the number...
3.9K
Surveys
14.8K
Often, psychologists develop surveys as a means of gathering data. Surveys are lists of questions to be answered by research participants, and can be delivered as paper-and-pencil questionnaires, administered electronically, or conducted verbally. Generally, the survey itself can be completed in a short time, and the ease of administering a survey makes it easy to collect data from a large number of people.
14.8K
