亲爱的研友该休息了!由于当前在线用户较少,发布求助请尽量完整地填写文献信息,科研通机器人24小时在线,伴您度过漫漫科研夜!身体可是革命的本钱,早点休息,好梦!

Assessing Ability for ChatGPT to Answer Total Knee Arthroplasty-Related Questions

关节置换术 清晰 骨科手术 相关性(法律) 一致性(知识库) 物理疗法 医学 外科 人工智能 计算机科学 生物化学 化学 政治学 法学
作者
Matthew L. Magruder,Ariel N. Rodriguez,Jason Wong,Orry Erez,Nicolás S. Piuzzi,Gil R. Scuderi,James Slover,Jason H. Oh,Ran Schwarzkopf,Antonia F. Chen,Richard Iorio,Stuart B. Goodman,Michael A. Mont
出处
期刊:Journal of Arthroplasty [Elsevier BV]
卷期号:39 (8): 2022-2027 被引量:23
标识
DOI:10.1016/j.arth.2024.02.023
摘要

Introduction Artificial intelligence (AI) in the field of orthopaedics has been a topic of increasing interest and opportunity in recent years. Its applications are widespread both for physicians and patients, including use in clinical decision-making, in the operating room, and in research. In this study, we aimed to assess the quality of ChatGPT answers when asked questions related to total knee arthroplasty (TKA). Methods ChatGPT prompts were created by turning 15 of the American Academy of Orthopaedic Surgeons (AAOS) Clinical Practice Guidelines into questions. An online survey was created, which included screenshots of each prompt and answers to the 15 questions. Surgeons were asked to grade ChatGPT answers from 1 to 5 based on their characteristics: 1) Relevance; 2) Accuracy; 3) Clarity; 4) Completeness; 5) Evidence-based; and 6) Consistency. There were eleven Adult Joint Reconstruction fellowship-trained surgeons who completed the survey. Questions were subclassified based on the subject of the prompt: 1) risk factors, 2) implant/Intraoperative, and 3) pain/functional outcomes. The average and standard deviation for all answers, as well as for each subgroup, were calculated. Inter-rater reliability (IRR) was also calculated. Results All answer characteristics were graded as being above average (i.e., a score > 3). Relevance demonstrated the highest scores (4.43±0.77) by surgeons surveyed, and consistency demonstrated the lowest scores (3.54±1.10). ChatGPT prompts in the Risk Factors group demonstrated the best responses, while those in the Pain/Functional Outcome group demonstrated the lowest. The overall IRR was found to be 0.33 (poor reliability), with the highest IRR for relevance (0.43) and the lowest for evidence-based (0.28). Conclusion ChatGPT can answer questions regarding well-established clinical guidelines in TKA with above-average accuracy but demonstrates variable reliability. This investigation is the first step in understanding large language model (LLM) AIs like ChatGPT and how well they perform in the field of arthroplasty.
最长约 10秒,即可获得该文献文件

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
24秒前
魁梧的天佑完成签到,获得积分10
28秒前
石愚志完成签到,获得积分10
32秒前
34秒前
健忘香彤完成签到,获得积分10
50秒前
58秒前
1分钟前
时尚沅完成签到,获得积分10
1分钟前
光亮豌豆完成签到,获得积分10
1分钟前
1分钟前
1分钟前
2分钟前
幸福的盼芙完成签到,获得积分10
2分钟前
wada3n完成签到,获得积分10
2分钟前
2分钟前
Criminology34应助KKSun采纳,获得10
2分钟前
美好的初翠完成签到,获得积分10
2分钟前
sailingluwl完成签到,获得积分10
2分钟前
3分钟前
自然谷波完成签到,获得积分10
3分钟前
可耐的萤完成签到,获得积分10
3分钟前
3分钟前
matrixu完成签到,获得积分10
3分钟前
Criminology34应助帅气的寄瑶采纳,获得30
3分钟前
小巧的傲晴完成签到,获得积分10
3分钟前
3分钟前
梅梅美美发布了新的文献求助10
4分钟前
兮豫完成签到 ,获得积分10
4分钟前
辛勤的涵菡完成签到,获得积分10
4分钟前
4分钟前
华仔应助干净的纸飞机采纳,获得10
4分钟前
4分钟前
4分钟前
漂亮的宛筠完成签到,获得积分10
4分钟前
活力的尔蓉完成签到,获得积分10
4分钟前
甜蜜寻琴完成签到,获得积分10
5分钟前
5分钟前
平淡大船完成签到,获得积分10
5分钟前
迷路的身影完成签到,获得积分10
5分钟前
梅梅美美完成签到,获得积分10
5分钟前
高分求助中
(应助此贴封号)【重要!!请各用户(尤其是新用户)详细阅读】【科研通的精品贴汇总】 10000
Principles of town planning: translating concepts to applications 1000
Sleep in the pediatric ICU: an empirical investigation 516
Management and the Arts 510
Matrix Methods in Data Mining and Pattern Recognition Second Edition 510
The Great Hymn to Šamaš 500
Positive Obsession: The Life and Times of Octavia E. Butler 500
热门求助领域 (近24小时)
化学 材料科学 医学 生物 纳米技术 工程类 有机化学 化学工程 生物化学 计算机科学 内科学 物理 复合材料 催化作用 细胞生物学 无机化学 光电子学 物理化学 电极 基因
热门帖子
关注 科研通微信公众号,转发送积分 7694114
求助须知:如何正确求助?哪些是违规求助? 9254693
关于积分的说明 19991105
捐赠科研通 7267768
什么是DOI,文献DOI怎么找? 3292006
关于科研通互助平台的介绍 2447934
邀请新用户注册赠送积分活动 2297423