可靠性(半导体)
可靠性工程
中医药
医学
内科学
计算机科学
遗传学
替代医学
生物
工程类
病理
物理
量子力学
功率(物理)
作者
Yun-Chi Wu,Yao‐Cheng Wu,Chunlin Wu,Tzuo-Yi Hsieh,Wen‐Wei Sung
标识
DOI:10.20944/preprints202501.1787.v1
摘要
The aim of our research was to evaluate the accuracies of different versions of ChatGPT in traditional Chinese medicine (TCM). We tested three versions of ChatGPT—GPT-3.5, GPT-4, and GPT-4o—using 960 questions from the first stage of the Professional and Technical Senior Examination for Doctors of Chinese Medicine in Taiwan. We found that only GPT-4o passed the exam, with an overall accuracy of 62.29%. Moreover, GPT-4 and GPT-4o performed better in Basic Chinese Medicine II (BCM II) than in Basic Chinese Medicine I (BCM I). We conclude that the GPT models demonstrate a stronger grasp of knowledge related to Chinese herbal formulas and Chinese materia medica in BCM II compared to their understanding of the history, basic theories, Neijing, and Nanjing in BCM I. Furthermore, a noticeable performance gap was evident between TCM and Western medicine. Because of the language bias in ChatGPT’s training on English datasets for TCM-related knowledge, more training is required with TCM-related Chinese data, especially in interpreting classical Chinese. Therefore, future research and development should further optimize the model’s performance in multilingual environments to advance the application of AI in medical education.
科研通智能强力驱动
Strongly Powered by AbleSci AI