2025 Expert consensus on retrospective evaluation of large language model applications in clinical scenarios

可比性 软件部署 计算机科学 医疗保健 钥匙(锁) 知识管理 数据科学 管理科学 自然语言 大数据 风险分析(工程) 医疗保健系统 专家系统 过程管理 人工智能
作者
Qing Chang,Fei Chen,Yaolong Chen,Longlong Cheng,Di Dong,Jiahong Dong,Xiaobin Feng,Junbo Ge,Jingjing He,Yihua He,Zhiyang He,Hong Ji,Xue Jiang,Zehua Jiang,Nan Li,Peng Li,Yazi Li,Bing Liu,Junwei Liu,Han Lyu
出处
期刊:Intelligent medicine [Elsevier]
卷期号:5 (4): 318-330 被引量:1
标识
DOI:10.1016/j.imed.2025.09.001
摘要

Large Language Models (LLMs), trained on vast amounts of textual data, have demonstrated strong capabilities in natural language understanding and generation. In the medical field, LLMs are increasingly applied across various domains such as disease screening, diagnostic assistance, and health management, playing a key role in advancing intelligent healthcare. In recent years, China has actively promoted the integration of artificial intelligence with healthcare through a series of policies that support enterprises in making breakthroughs in key technologies such as medical large language models and multi-modal data integration. Concurrently, efforts have accelerated the deployment of AI in applications such as health management and precision medicine, to gradually establish a full-cycle intelligent healthcare system encompassing prevention, diagnosis, treatment, and rehabilitation. However, the rapid deployment of LLMs in healthcare has highlighted the lack of standardized evaluation criteria and consistent methodologies. To address this, this expert consensus focuses on establishing a retrospective evaluation framework tailored to medical applications. By integrating scientific evaluation metrics, standards, and procedures, the framework provides clear and actionable guidance for model evaluators, developers, and end-users. It aims to unify assessment practices, enhance the scientific rigor and comparability of evaluations, and ensure the safe and effective use of LLMs in healthcare, ultimately supporting the high-quality development of AI-powered medical services.
最长约 10秒,即可获得该文献文件

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
跳跃的鹏飞完成签到 ,获得积分0
7秒前
赵亦恬完成签到 ,获得积分10
16秒前
PP的应助被寒冷的断秋采纳,获得10
19秒前
22秒前
白雪完成签到,获得积分10
23秒前
美丽人生完成签到 ,获得积分10
23秒前
寒冷的断秋完成签到,获得积分10
24秒前
一投就中完成签到 ,获得积分10
31秒前
CJH完成签到,获得积分0
38秒前
38秒前
jiyixiyang完成签到,获得积分10
47秒前
49秒前
kitsch完成签到 ,获得积分10
51秒前
淡定煎饼完成签到,获得积分10
53秒前
留胡子的寄瑶完成签到,获得积分10
55秒前
狗狼狼完成签到,获得积分10
56秒前
义气绍辉完成签到,获得积分10
59秒前
趙途嘵生完成签到,获得积分10
59秒前
迷人的语芹完成签到,获得积分10
1分钟前
1分钟前
辰辰完成签到 ,获得积分10
1分钟前
郑浩完成签到,获得积分10
1分钟前
1分钟前
ERIC完成签到,获得积分10
1分钟前
1分钟前
英姑的应助被寒冷的断秋采纳,获得30
1分钟前
planto完成签到,获得积分10
1分钟前
七安完成签到 ,获得积分10
1分钟前
lilac完成签到 ,获得积分10
1分钟前
xiaolizi完成签到,获得积分0
1分钟前
Polylactic完成签到 ,获得积分10
1分钟前
大力的夜玉完成签到 ,获得积分10
1分钟前
呆橘完成签到 ,获得积分10
1分钟前
急聘行完成签到 ,获得积分10
1分钟前
1分钟前
123啊完成签到 ,获得积分10
1分钟前
尊敬康乃馨完成签到,获得积分10
1分钟前
晨晨完成签到 ,获得积分10
1分钟前
菠萝吹雪完成签到,获得积分10
1分钟前
林林爱学医完成签到 ,获得积分10
1分钟前
高分求助中
(应助此贴封号)通过应助OA文献获取积分 10000
Rosenblum, Global Change Biology 800
The Dawn of Philology 520
Organizational Behavior 510
Production Logging: Theoretical and Interpretive Elements 400
A primer on partial least squares structural equation modeling (PLS-SEM) (4th ed.) 310
中国器官捐献和移植发展报告(2024) 300
热门求助领域 (近24小时)
化学 材料科学 医学 生物 计算机科学 工程类 纳米技术 内科学 物理 有机化学 化学工程 生物化学 复合材料 光电子学 细胞生物学 心理学 量子力学 催化作用 物理化学 电极
热门帖子
关注 科研通微信公众号,转发送积分 7820416
求助须知:如何正确求助?哪些是违规求助? 9347793
关于积分的说明 20543187
捐赠科研通 7413030
什么是DOI,文献DOI怎么找? 3332686
关于科研通互助平台的介绍 2478638
邀请新用户注册赠送积分活动 2352732