How AI responds to common HIV/AIDS questions: ChatGPT versus DeepSeek

答疑 人工智能 自然语言处理 传染病(医学专业) 显著性差异 金标准(测试) 心理学 比例(比率) 计算机科学 机器学习 医学 梅德林 F1得分 情报检索 平均差 认知心理学 语言模型 数据科学 疾病
作者
Hui Huang,Huichao Zhang,Xinxin Qin,Yuhan Wu,Qi Yu,Mingxia Fang,Binghu Sun,Jian Cheng,Yan Song,Junmei Shu,Ling Wang
出处
期刊:Frontiers in Public Health [Frontiers Media]
卷期号:14: 1886493-1886493
标识
DOI:10.3389/fpubh.2026.1886493
摘要

Background Artificial intelligence (AI) has garnered significant attention and, to some extent, has even replaced certain search engines as a popular channel for people across regions to access information. HIV/AIDS is a common topic among infectious diseases, and patients or high-risk groups frequently search for information online. Our study evaluated the accuracy of two AI models (ChatGPT-3.5 and DeepSeek-R1) in answering HIV/AIDS-related questions. Methods This was a cross-sectional analytical study comparing two advanced large language models representing Eastern and Western cultural backgrounds, respectively, with a gold standard (an infectious disease specialist). We compiled 26 common questions related to HIV/AIDS and categorized them into four themes (basic knowledge, diagnosis, treatment, and prevention). Three infectious disease experts independently scored the AI models’ responses on a 4-point scale for each answer. There was no statistically significant difference in the scores for HIV/AIDS-related questions between the two AI models: Z = −2.135, two-sided p -value = 0.451 ( p > 0.05). DeepSeek-R1 demonstrated higher accuracy, with 53.85% of its responses rated as “Excellent,” compared to 48.72% for ChatGPT-3.5; there was no significant difference in accuracy between the two AI models when answering HIV/AIDS-related questions ( χ 2 = 0.429, df = 1, p = 0.513). Both AI models achieved high average composite scores (DeepSeek-R1 = 3.50, ChatGPT-3.5 = 3.44, out of a maximum of 4 points). The AI models performed exceptionally well across all domains, though they were relatively weaker in the “Treatment” domain. Conclusion Our study has identified the potential of AI models, particularly ChatGPT-3.5 and DeepSeek-R1, to provide accurate and comprehensive responses to HIV/AIDS-related questions. As free online tools, they can provide personalized, useful medical information to diverse populations across regions. However, medical information generated by AI models should still be used under the supervision and review of healthcare professionals.
最长约 10秒,即可获得该文献文件

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
直率小之完成签到,获得积分20
刚刚
共享精神应助renyi采纳,获得10
1秒前
biochen发布了新的文献求助10
1秒前
大地发布了新的文献求助10
1秒前
科研通AI6.4应助张嚼采纳,获得10
2秒前
NexusExplorer应助fdj3121采纳,获得10
2秒前
2秒前
裴瑞志发布了新的文献求助10
3秒前
3秒前
xyj完成签到,获得积分10
3秒前
4秒前
绵绵发布了新的文献求助10
4秒前
Akim应助hooo采纳,获得10
5秒前
5秒前
三火发布了新的文献求助10
7秒前
hanjian完成签到,获得积分10
7秒前
和气生财君完成签到 ,获得积分0
7秒前
清爽凝安发布了新的文献求助10
8秒前
lu完成签到,获得积分10
8秒前
隐形曼青应助周沁圆采纳,获得10
8秒前
9秒前
凡松应助裴瑞志采纳,获得10
9秒前
9秒前
13508104971发布了新的文献求助10
10秒前
10秒前
风清扬发布了新的文献求助10
11秒前
12秒前
xiaofan完成签到,获得积分10
12秒前
满意的起眸完成签到,获得积分10
13秒前
大地完成签到,获得积分10
14秒前
renyi发布了新的文献求助10
15秒前
科研小白发布了新的文献求助10
15秒前
华仔应助科研通管家采纳,获得10
17秒前
molihuakai应助科研通管家采纳,获得10
17秒前
NexusExplorer应助科研通管家采纳,获得10
17秒前
Lucas应助科研通管家采纳,获得10
17秒前
17秒前
烟花应助科研通管家采纳,获得10
17秒前
17秒前
隐形曼青应助科研通管家采纳,获得10
17秒前
高分求助中
(应助此贴封号)【重要!!请各用户(尤其是新用户)详细阅读】【科研通的精品贴汇总】 10000
Industrial Hydraulics Manual (7th edition) 800
Physiologic races of the downy mildew fungus on soybeans in North Carolina 800
Rosenblum, Global Change Biology 800
Essentials of Carbohydrate Chemistry and Biochemistry, 4th Edition 800
Organizational Behavior 510
Management and the Arts 510
热门求助领域 (近24小时)
化学 材料科学 医学 生物 纳米技术 计算机科学 化学工程 工程类 有机化学 物理 复合材料 生物化学 内科学 细胞生物学 基因 遗传学 免疫学 冶金 光电子学 癌症研究
热门帖子
关注 科研通微信公众号,转发送积分 7775985
求助须知:如何正确求助?哪些是违规求助? 9317495
关于积分的说明 20357869
捐赠科研通 7362388
什么是DOI,文献DOI怎么找? 3318104
关于科研通互助平台的介绍 2466309
邀请新用户注册赠送积分活动 2333431