Few‐Shot Lung Cancer Classification via Electronic Nose Using Large Language Models: A Multicentre Prospective Study

医学 肺癌 卷积神经网络 人工智能 灵敏度(控制系统) 前瞻性队列研究 癌症 电子鼻 放射科 适应(眼睛) 鼻子 诊断准确性 医学物理学 班级(哲学) 机器学习
作者
Meng‐Rui Lee,Chun‐Yao Huang,Joyce Yue Sun,Chang‐Ru Lin,Wen‐Yuan Lin,Nai‐Hui Chi,Kea‐Tiong Tang,Jann‐Yuan Wang,Chao‐Chi Ho,Jin‐Yuan Shih,Chong‐Jen Yu
出处
期刊:Respirology [Wiley]
被引量:1
标识
DOI:10.1002/resp.70266
摘要

BACKGROUND AND OBJECTIVE: Electronic Nose (eNose) breathprints are promising non-invasive lung cancer diagnostic tools, but cross-site validation and adaptation remain barriers to clinical applications. It remains unknown whether a natural language processing-pretrained large language model (LLM) can enable few-shot, site-specific classification of lung cancer using eNose breathprints. METHODS: We collected eNose breathprints of lung cancer and non-lung cancer patients from two medical centres in Taiwan. A GPT-2-backbone LLM with parameter-efficient adaptation was compared with convolutional neural networks (CNN) trained from scratch or pretrained on CIFAR-100. Few-shot protocols (2-6 shots per class) and full-data training were evaluated. RESULTS: We collected 432 eNose breathprints from two sites (S1 and S2). With 6 labelled samples per class (6 shots), LLM achieved an area under the curve (AUC) of 0.79 (95% CI: 0.71-0.87), sensitivity of 0.74 (0.63-0.83), and specificity of 0.77 (0.67-0.87) on S1. On S2, it achieved an AUC of 0.76 (0.69-0.82), sensitivity of 0.77 (0.69-0.84), and specificity of 0.61 (0.51-0.70). LLM outperforms scratch CNN models (S1; AUC: 0.44, p = 0.0002) (S2; AUC: 0.63, p = 0.0198) and CNN pretrained on CIFAR-100 images (S1; AUC: 0.57, p = 0.0100) and (S2; AUC: 0.61, p = 0.0248). LLM or a CNN model trained on the source site fails to improve performance after transferring to the target site for fine-tuning; for the LLM, performance even deteriorates. CONCLUSION: Our study demonstrates the potential of pretrained LLMs for few-shot lung cancer classification in a real-world mixed clinical cohort, reducing dependence on large training datasets.
最长约 10秒,即可获得该文献文件

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
jijibao发布了新的文献求助10
1秒前
ilc发布了新的文献求助10
1秒前
dachang完成签到,获得积分10
1秒前
2秒前
火柴盒完成签到,获得积分10
3秒前
way_oz发布了新的文献求助10
3秒前
3秒前
dachang发布了新的文献求助10
4秒前
4秒前
5秒前
冷静的口红完成签到,获得积分10
5秒前
烟花应助九花玉露丸采纳,获得10
6秒前
MM完成签到,获得积分10
6秒前
7秒前
尉迟莲发布了新的文献求助10
7秒前
传奇3应助纪清月采纳,获得10
8秒前
8秒前
8秒前
大方岩完成签到,获得积分10
8秒前
隐形曼青应助甜美的寒珊采纳,获得10
9秒前
小蘑菇应助专注的妙竹采纳,获得10
9秒前
9秒前
11秒前
你在烦恼什么完成签到,获得积分10
11秒前
香蕉觅云应助Domenico采纳,获得10
12秒前
zzt发布了新的文献求助10
13秒前
银鱼在游发布了新的文献求助10
13秒前
Owen应助可可豆战士采纳,获得10
13秒前
是人完成签到 ,获得积分10
14秒前
lzc4632发布了新的文献求助20
14秒前
QQ完成签到 ,获得积分10
15秒前
完美世界应助欢呼的道之采纳,获得10
15秒前
15秒前
XyuF完成签到,获得积分10
16秒前
阿鑫发布了新的文献求助10
16秒前
16秒前
16秒前
张天完成签到,获得积分10
17秒前
Once发布了新的文献求助10
18秒前
18秒前
高分求助中
(应助此贴封号)【重要!!请各用户(尤其是新用户)详细阅读】【科研通的精品贴汇总】 10000
China Pluperfect I: Epistemology of Past and Outside in Chinese Art 520
Matrix Methods in Data Mining and Pattern Recognition Second Edition 510
Governing Growth: Us Industrial Policy from Hamilton to Trump 500
The fast track to determining transfer functions of linear circuits: The student guide 500
The Analytical and Numerical Solution of Electric and Magnetic Fields 500
Synthesis of P-Chiral Phosphine Ligands and Their Applications in Asymmetric Catalysis 400
热门求助领域 (近24小时)
化学 材料科学 医学 生物 纳米技术 工程类 有机化学 化学工程 生物化学 计算机科学 内科学 物理 复合材料 催化作用 细胞生物学 无机化学 光电子学 物理化学 电极 基因
热门帖子
关注 科研通微信公众号,转发送积分 7623982
求助须知:如何正确求助?哪些是违规求助? 9199123
关于积分的说明 19721838
捐赠科研通 7195185
什么是DOI,文献DOI怎么找? 3273428
关于科研通互助平台的介绍 2435587
邀请新用户注册赠送积分活动 2269167