亲爱的研友该休息了!由于当前在线用户较少,发布求助请尽量完整地填写文献信息,科研通机器人24小时在线,伴您度过漫漫科研夜!身体可是革命的本钱,早点休息,好梦!

Large Language Models for Diagnosing Focal Liver Lesions From CT/MRI Reports: A Comparative Study With Radiologists

医学诊断 组织病理学 回顾性队列研究 医学 鉴别诊断 放射科 磁共振成像 病理
作者
Liuji Sheng,Yidi Chen,Hong Wei,Feng Che,Yingyi Wu,Qin Qin,Chongtu Yang,Yanshu Wang,Jingwen Peng,Mustafa R. Bashir,Maxime Ronot,Bin Song,Hanyu Jiang
出处
期刊:Liver International [Wiley]
卷期号:45 (6): e70115-e70115 被引量:15
标识
DOI:10.1111/liv.70115
摘要

BACKGROUND & AIMS: Whether large language models (LLMs) could be integrated into the diagnostic workflow of focal liver lesions (FLLs) remains unclear. We aimed to investigate two generic LLMs (ChatGPT-4o and Gemini) regarding their diagnostic accuracies referring to the CT/MRI reports, compared to and combined with radiologists of different experience levels. METHODS: From April 2022 to April 2024, this single-center retrospective study included consecutive adult patients who underwent contrast-enhanced CT/MRI for single FLL and subsequent histopathologic examination. The LLMs were prompted by clinical information and the "findings" section of radiology reports three times to provide differential diagnoses in the descending order of likelihood, with the first considered the final diagnosis. In the research setting, six radiologists (three junior and three middle-level) independently reviewed the CT/MRI images and clinical information in two rounds (first alone, then with LLM assistance). In the clinical setting, diagnoses were retrieved from the "impressions" section of radiology reports. Diagnostic accuracy was investigated against histopathology. RESULTS: 228 patients (median age, 59 years; 155 males) with 228 FLLs (median size, 3.6 cm) were included. Regarding the final diagnosis, the accuracy of two-step ChatGPT-4o (78.9%) was higher than single-step ChatGPT-4o (68.0%, p < 0.001) and single-step Gemini (73.2%, p = 0.004), similar to real-world radiology reports (80.0%, p = 0.34) and junior radiologists (78.9%-82.0%; p-values, 0.21 to > 0.99), but lower than middle-level radiologists (84.6%-85.5%; p-values, 0.001 to 0.02). No incremental diagnostic value of ChatGPT-4o was observed for any radiologist (p-values, 0.63 to > 0.99). CONCLUSION: Two-step ChatGPT-4o showed matching accuracies to real-world radiology reports and junior radiologists for diagnosing FLLs but was less accurate than middle-level radiologists and demonstrated little incremental diagnostic value.
最长约 10秒,即可获得该文献文件

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
1秒前
完美世界应助发文章12138采纳,获得10
2秒前
7秒前
Septvirouo9完成签到,获得积分10
13秒前
激动的水蓝完成签到,获得积分10
14秒前
19秒前
yy完成签到 ,获得积分10
19秒前
23秒前
科研通AI6.4应助ivy1991采纳,获得10
24秒前
redeem发布了新的文献求助10
27秒前
科研通AI6.2应助Septvirouo9采纳,获得10
28秒前
科研通AI2S应助发文章12138采纳,获得80
29秒前
30秒前
jhbwudiwudi完成签到,获得积分10
31秒前
星辰大海应助su采纳,获得10
32秒前
34秒前
37秒前
LUCA完成签到 ,获得积分10
37秒前
cdercder应助ZHJ采纳,获得10
38秒前
40秒前
41秒前
简珹楚完成签到 ,获得积分10
42秒前
阔达的芹菜完成签到,获得积分10
43秒前
无极微光应助沉默的涔采纳,获得20
43秒前
打打应助发文章12138采纳,获得10
47秒前
大个应助犹豫的大碗采纳,获得10
54秒前
小菀儿完成签到 ,获得积分10
56秒前
redeem完成签到,获得积分10
57秒前
cdercder应助ZHJ采纳,获得10
1分钟前
生动的孤容完成签到 ,获得积分10
1分钟前
1分钟前
1分钟前
狂野冬寒完成签到,获得积分10
1分钟前
清爽水之完成签到,获得积分10
1分钟前
打打应助盒盒怪采纳,获得10
1分钟前
酷波er应助盒盒怪采纳,获得10
1分钟前
kjjjj发布了新的文献求助10
1分钟前
彭于晏应助盒盒怪采纳,获得10
1分钟前
慕青应助盒盒怪采纳,获得10
1分钟前
打打应助盒盒怪采纳,获得10
1分钟前
高分求助中
(应助此贴封号)【重要!!请各用户(尤其是新用户)详细阅读】【科研通的精品贴汇总】 10000
Principles of town planning: translating concepts to applications 1000
内視鏡的に摘除しえた十二指腸乳頭部腫瘍の2例 660
Management and the Arts 510
Matrix Methods in Data Mining and Pattern Recognition Second Edition 510
Positive Obsession: The Life and Times of Octavia E. Butler 500
Interpolation and Regression Models for the Chemical Engineer: Solving Numerical Problems 400
热门求助领域 (近24小时)
化学 材料科学 医学 生物 纳米技术 工程类 有机化学 化学工程 生物化学 计算机科学 内科学 物理 复合材料 催化作用 细胞生物学 无机化学 光电子学 物理化学 电极 基因
热门帖子
关注 科研通微信公众号,转发送积分 7687668
求助须知:如何正确求助?哪些是违规求助? 9250588
关于积分的说明 19963645
捐赠科研通 7260646
什么是DOI,文献DOI怎么找? 3289878
关于科研通互助平台的介绍 2446781
邀请新用户注册赠送积分活动 2294522