亲爱的研友该休息了!由于当前在线用户较少,发布求助请尽量完整地填写文献信息,科研通机器人24小时在线,伴您度过漫漫科研夜!身体可是革命的本钱,早点休息,好梦!

Evaluating the Evolution of ChatGPT as an Information Resource in Shoulder and Elbow Surgery

医学 肘部 外科 骨科手术
作者
Benjamin Nieves-Lopez,Alexandra Bechtle,Jennifer Traverse,Christopher S. Klifto,Bradley S. Schoch,Keith T. Aziz
出处
期刊:Orthopedics [Slack Incorporated (United States)]
卷期号:: 1-6 被引量:3
标识
DOI:10.3928/01477447-20250123-03
摘要

The purpose of this study was to evaluate the performance and evolution of Chat Generative Pre-Trained Transformer (ChatGPT; OpenAI) as a resource for shoulder and elbow surgery information by assessing its accuracy on the American Academy of Orthopaedic Surgeons shoulder-elbow self-assessment questions. We hypothesized that both ChatGPT models would demonstrate proficiency and that there would be significant improvement with progressive iterations. A total of 200 questions were selected from the 2019 and 2021 American Academy of Orthopaedic Surgeons shoulder-elbow self-assessment questions. ChatGPT 3.5 and 4 were used to evaluate all questions. Questions with non-text data were excluded (114 questions). Remaining questions were input into ChatGPT and categorized as follows: anatomy, arthroplasty, basic science, instability, miscellaneous, nonoperative, and trauma. ChatGPT's performances were quantified and compared across categories with chi-square tests. The continuing medical education credit threshold of 50% was used to determine proficiency. Statistical significance was set at P<.05. ChatGPT 3.5 and 4 answered 52.3% and 73.3% of the questions correctly, respectively (P=.003). ChatGPT 3.5 performed significantly better in the instability category (P=.037). ChatGPT 4's performance did not significantly differ across categories (P=.841). ChatGPT 4 performed significantly better than ChatGPT 3.5 in all categories except instability and miscellaneous. ChatGPT 3.5 and 4 exceeded the proficiency threshold. ChatGPT 4 performed better than ChatGPT 3.5, showing an increased capability to correctly answer shoulder and elbow-focused questions. Further refinement of ChatGPT's training may improve its performance and utility as a resource. Currently, ChatGPT remains unable to answer questions at a high enough accuracy to replace clinical decision-making. [Orthopedics. 202x;4x(x):xx-xx.].

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
4秒前
勤劳半芹发布了新的文献求助10
5秒前
欣慰怀梦完成签到,获得积分10
5秒前
6秒前
彭于晏应助李斌采纳,获得10
7秒前
8秒前
9秒前
sukasuka发布了新的文献求助10
12秒前
Orange应助勤劳半芹采纳,获得10
13秒前
shuishui发布了新的文献求助10
14秒前
十一完成签到 ,获得积分10
15秒前
16秒前
shuishui完成签到,获得积分10
23秒前
24秒前
苗条的采梦完成签到,获得积分10
25秒前
27秒前
美满一寡完成签到,获得积分10
30秒前
sukasuka发布了新的文献求助10
31秒前
aa发布了新的文献求助10
31秒前
33秒前
35秒前
田様应助科研通管家采纳,获得10
35秒前
35秒前
35秒前
JamesPei应助sukasuka采纳,获得10
40秒前
怪不好意思的完成签到 ,获得积分10
42秒前
谦让的忆枫完成签到,获得积分10
54秒前
56秒前
59秒前
拼搏的水桃完成签到,获得积分10
59秒前
sukasuka发布了新的文献求助10
1分钟前
aa完成签到,获得积分10
1分钟前
小二郎应助2589采纳,获得10
1分钟前
1分钟前
顾矜应助sukasuka采纳,获得10
1分钟前
13074758911发布了新的文献求助10
1分钟前
1分钟前
1分钟前
molihuakai应助神奇的大蛇丸采纳,获得10
1分钟前
jiang完成签到,获得积分10
1分钟前
高分求助中
(应助此贴封号)【重要!!请各用户(尤其是新用户)详细阅读】【科研通的精品贴汇总】 10000
Effects of Two Weeks of Red Light Therapy on Choroidal Thickness and Axial Length in Young Adults 700
Management and the Arts 510
Matrix Methods in Data Mining and Pattern Recognition Second Edition 510
The Neuroscience of Language 400
Common Foundations of American and East Asian Modernisation: From Alexander Hamilton to Junichero Koizumi 400
Too Much of Two Good Things: Investment Protection and Environmental Protection in International Law 260
热门求助领域 (近24小时)
化学 材料科学 医学 生物 纳米技术 工程类 有机化学 化学工程 生物化学 计算机科学 内科学 物理 复合材料 催化作用 细胞生物学 无机化学 光电子学 物理化学 电极 基因
热门帖子
关注 科研通微信公众号,转发送积分 7673238
求助须知:如何正确求助?哪些是违规求助? 9239888
关于积分的说明 19902711
捐赠科研通 7242715
什么是DOI,文献DOI怎么找? 3285537
关于科研通互助平台的介绍 2443564
邀请新用户注册赠送积分活动 2287759