Evaluating the Effectiveness of Large Language Models in Providing Patient Education for Chinese Patients With Ocular Myasthenia Gravis: Mixed Methods Study

可读性 可用性 医疗保健 医学 重症肌无力 家庭医学 病人教育 复视 医学教育 心理学 计算机科学 外科 内科学 人机交互 经济 经济增长 程序设计语言
作者
Bin Wei,Yao Lian,Xin Hu,Y.C. Hu,Jie Rao,Yu Ji,Zhiyong Dong,Yichong Duan,Xiaorong Wu
出处
期刊:Journal of Medical Internet Research [JMIR Publications]
卷期号:27: e67883-e67883 被引量:2
标识
DOI:10.2196/67883
摘要

Background Ocular myasthenia gravis (OMG) is a neuromuscular disorder primarily affecting the extraocular muscles, leading to ptosis and diplopia. Effective patient education is crucial for disease management; however, in China, limited health care resources often restrict patients’ access to personalized medical guidance. Large language models (LLMs) have emerged as potential tools to bridge this gap by providing instant, AI-driven health information. However, their accuracy and readability in educating patients with OMG remain uncertain. Objective The purpose of this study was to systematically evaluate the effectiveness of multiple LLMs in the education of Chinese patients with OMG. Specifically, the validity of these models in answering patients with OMG-related questions was assessed through accuracy, completeness, readability, usefulness, and safety, and patients’ ratings of their usability and readability were analyzed. Methods The study was conducted in two phases: 130 choice ophthalmology examination questions were input into 5 different LLMs. Their performance was compared with that of undergraduates, master’s students, and ophthalmology residents. In addition, 23 common patients with OMG-related patient questions were posed to 4 LLMs, and their responses were evaluated by ophthalmologists across 5 domains. In the second phase, 20 patients with OMG interacted with the 2 LLMs from the first phase, each asking 3 questions. Patients assessed the responses for satisfaction and readability, while ophthalmologists evaluated the responses again using the 5 domains. Results ChatGPT o1-preview achieved the highest accuracy rate of 73% on 130 ophthalmology examination questions, outperforming other LLMs and professional groups like undergraduates and master’s students. For 23 common patients with OMG-related questions, ChatGPT o1-preview scored highest in correctness (4.44), completeness (4.44), helpfulness (4.47), and safety (4.6). GEMINI (Google DeepMind) provided the easiest-to-understand responses in readability assessments, while GPT-4o had the most complex responses, suitable for readers with higher education levels. In the second phase with 20 patients with OMG, ChatGPT o1-preview received higher satisfaction scores than Ernie 3.5 (Baidu; 4.40 vs 3.89, P=.002), although Ernie 3.5’s responses were slightly more readable (4.31 vs 4.03, P=.01). Conclusions LLMs such as ChatGPT o1-preview may have the potential to enhance patient education. Addressing challenges such as misinformation risk, readability issues, and ethical considerations is crucial for their effective and safe integration into clinical practice.
最长约 10秒,即可获得该文献文件

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
珺儿发布了新的文献求助10
刚刚
刚刚
Owen应助科研通管家采纳,获得10
刚刚
彭于晏应助顺利汉堡采纳,获得30
刚刚
刚刚
充电宝应助科研通管家采纳,获得10
刚刚
无花果应助科研通管家采纳,获得10
刚刚
刚刚
xhj666完成签到,获得积分10
刚刚
科研通AI2S应助zzx采纳,获得10
刚刚
酷波er应助科研通管家采纳,获得10
刚刚
思源应助科研通管家采纳,获得30
1秒前
脑洞疼应助超帅的黄采纳,获得10
1秒前
爆米花应助科研通管家采纳,获得10
1秒前
orixero应助科研通管家采纳,获得10
1秒前
知诵发布了新的文献求助10
1秒前
田洪涛发布了新的文献求助10
1秒前
852应助科研通管家采纳,获得10
1秒前
1秒前
慕青应助科研通管家采纳,获得10
1秒前
1秒前
完美世界应助逢春采纳,获得10
1秒前
李健应助科研通管家采纳,获得10
2秒前
张正发布了新的文献求助10
2秒前
aali完成签到,获得积分10
2秒前
田様应助科研通管家采纳,获得10
2秒前
2秒前
2秒前
顺顺发布了新的文献求助10
2秒前
自信的冷雁完成签到 ,获得积分10
2秒前
CipherSage应助科研通管家采纳,获得10
2秒前
2秒前
2秒前
2秒前
3秒前
丘比特应助科研通管家采纳,获得10
3秒前
脑洞疼应助科研通管家采纳,获得10
3秒前
辉常奋斗完成签到,获得积分10
3秒前
打打应助科研通管家采纳,获得10
3秒前
Jelinna发布了新的文献求助10
3秒前
高分求助中
(应助此贴封号)【重要!!请各用户(尤其是新用户)详细阅读】【科研通的精品贴汇总】 10000
The anomeric effect 1314
Principles of town planning: translating concepts to applications 1000
1 Peter and Christ's Descent to the Dead in Its Early Christian Reception 700
Organizational Behavior 510
Management and the Arts 510
Matrix Methods in Data Mining and Pattern Recognition Second Edition 510
热门求助领域 (近24小时)
化学 材料科学 医学 生物 纳米技术 工程类 有机化学 化学工程 生物化学 计算机科学 内科学 物理 复合材料 催化作用 细胞生物学 无机化学 光电子学 物理化学 电极 基因
热门帖子
关注 科研通微信公众号,转发送积分 7735856
求助须知:如何正确求助?哪些是违规求助? 9285977
关于积分的说明 20174696
捐赠科研通 7314073
什么是DOI,文献DOI怎么找? 3305151
关于科研通互助平台的介绍 2457568
邀请新用户注册赠送积分活动 2314539