Evaluating a Chatbot as a Companion for Patients With Breast Cancer: Collaborative Pilot Study

预印本 聊天机器人 乳腺癌 医学 计算机科学 癌症 万维网 内科学
作者
Sebastian Daniel Boie,Esther Glastetter,Michael P. Lux,Felix Balzer,Christof von Kalle,Christian Lenz,Ulrich Müller
出处
期刊:JMIR cancer [JMIR Publications]
卷期号:11: e68426-e68426
标识
DOI:10.2196/68426
摘要

Abstract Background Patients with breast cancer frequently experience significant uncertainty, prompting them to seek detailed, personalized, and reliable medical information to enhance adherence to prescribed treatments, medications, and recommended lifestyle adjustments. Although high-quality information exists within oncology guidelines and patient-oriented resources, the provision of tailored responses to individual patient queries remains challenging, especially for non–English-speaking populations. Objective This study aims to evaluate the potential of an artificial intelligence–driven chatbot, specifically leveraging ChatGPT (GPT-4; OpenAI) combined with retrieval-augmented generation, to deliver personalized answers to complex breast cancer-related patient questions in German. Methods We collaborated with one of Germany’s largest breast cancer Patient Representation Groups to collect authentic patient inquiries, receiving a total of 118 questions. After initial screening, we selected 104 medical questions, organized into 7 distinct categories: aftercare, bone health, ductal carcinoma in situ, diagnostics, nutrition and supplements, complementary medicine, and therapy. A customized version of GPT-4 was configured with specific system prompts emphasizing empathetic, evidence-based responses and integrated with a comprehensive database comprising guidelines, recommendations, and patient information materials published by recognized German medical societies. To assess chatbot responses, we used 4 evaluation criteria: comprehensibility (clarity from a patient perspective), correctness (accuracy per current medical guidelines), completeness (inclusion of all relevant aspects), and potential harm (risk of undue patient harm or misinformation). Ratings were conducted using a 5-point Likert scale by a breast cancer expert (correctness, completeness, and potential harm) and patient representatives (comprehensibility). Results The chatbot provided high-quality responses across multiple dimensions. Of the 499 responses evaluated for comprehensibility, 427 (85.6%) were rated as comprehensible. Among the 104 responses assessed for the remaining dimensions, 91 (87.5%) were rated as correct, 72 (69.2%) as complete, and 93 (89.4%) as nonharmful. Reasons for incomplete answers included omission of reimbursement details, updates from recent therapeutic guidelines, or nuanced recommendations regarding endocrine therapy and aftercare schedules. In addition, 6 (5.8%) of the answers were rated as potentially harmful due to outdated or contextually inappropriate recommendations. The chatbot also performed well in the nutrition and bone health categories despite occasionally incomplete document retrieval. Conclusions Our findings demonstrate that an artificial intelligence–powered chatbot with GPT-4 and retrieval augmentation can effectively provide personalized, linguistically accessible, and largely accurate information to German-speaking patients with breast cancer. This approach holds considerable promise for improving patient-centered communication, empowering patients to make informed decisions. Nonetheless, observed limitations regarding response completeness and potential harm underscore the critical need for ongoing human oversight. Future research and development should prioritize regularly updated databases, advanced retrieval methods to handle complex document structures, multimodal capabilities, and clearly articulated disclaimers emphasizing the necessity of professional medical consultation. Our evaluation, along with the provided set of realistic patient questions, establishes a benchmark for future development and validation of German-language oncology chatbots.

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
yyyyyz完成签到,获得积分10
刚刚
初見发布了新的文献求助10
刚刚
小瘦猴完成签到,获得积分10
1秒前
1秒前
研友_8yN60L发布了新的文献求助10
1秒前
1秒前
1秒前
ling完成签到,获得积分10
1秒前
栀染完成签到,获得积分10
2秒前
天青妖发布了新的文献求助10
2秒前
尊敬凝荷发布了新的文献求助10
2秒前
ZHQ完成签到,获得积分10
3秒前
3秒前
含蓄含烟完成签到,获得积分10
3秒前
3秒前
背后幻竹完成签到,获得积分10
3秒前
3秒前
piggyfly完成签到,获得积分10
4秒前
4秒前
TTT完成签到,获得积分10
4秒前
未来完成签到,获得积分10
4秒前
科研通AI6.4的应助被KD采纳,获得10
5秒前
li完成签到,获得积分10
5秒前
CodeCraft的应助被黑鹿采纳,获得10
5秒前
CX完成签到,获得积分10
5秒前
大模型的应助被慧_h采纳,获得10
5秒前
壮观的冬云完成签到,获得积分10
5秒前
东北雨姐发布了新的文献求助10
6秒前
CipherSage的应助被xuzan采纳,获得10
6秒前
mafei完成签到,获得积分10
6秒前
Ava的应助被科研通管家采纳,获得10
6秒前
Nole的应助被科研通管家采纳,获得10
6秒前
热切菩萨的应助被科研通管家采纳,获得10
6秒前
torch132完成签到,获得积分0
6秒前
李治博发布了新的文献求助10
6秒前
DOC_XIONG的应助被科研通管家采纳,获得10
6秒前
无限尔曼完成签到 ,获得积分10
6秒前
泪了睡吧的应助被科研通管家采纳,获得10
7秒前
7秒前
大模型的应助被科研通管家采纳,获得10
7秒前
高分求助中
(应助此贴封号)【重要!!请各用户(尤其是新用户)详细阅读】【科研通的精品贴汇总】 10000
Rosenblum, Global Change Biology 800
Organizational Behavior 510
Management and the Arts 510
Geschichtliche Grundbegriffe (GGB), Band 5: Pro–Soz 300
Die Religion in Geschichte und Gegenwart (RGG), 4. Auflage, Band 7: R–S 300
Die Religion in Geschichte und Gegenwart (RGG), 4. Auflage, Band 1: A–B 300
热门求助领域 (近24小时)
化学 材料科学 医学 生物 计算机科学 工程类 纳米技术 内科学 物理 有机化学 化学工程 生物化学 复合材料 光电子学 细胞生物学 心理学 量子力学 催化作用 物理化学 电极
热门帖子
关注 科研通微信公众号,转发送积分 7792952
求助须知:如何正确求助?哪些是违规求助? 9329754
关于积分的说明 20431920
捐赠科研通 7382722
什么是DOI,文献DOI怎么找? 3323849
关于科研通互助平台的介绍 2471748
邀请新用户注册赠送积分活动 2340937