标杆管理
透明度(行为)
持续性
计算机科学
钥匙(锁)
管理科学
风险分析(工程)
宣传
数据科学
工程伦理学
水准点(测量)
可持续发展
临床决策
可读性
梅德林
医学
光学(聚焦)
作者
Xiaofei Wang,Zhuxin Xiong,Ke Zou,Sahana Srinivasan,Thaddaeus Wai Soon Lo,Yilan Wu,Minjie Zou,Nan Liu,Fares Antaki,Weizhi Ma,Seyed Mohammad Nabavi,Julian Savulescu,Josip Car,David C Klonoff,Bin Sheng,Tien Yin Wong,Qingyu Chen,Yih Chung Tham
标识
DOI:10.1016/j.landig.2025.100931
摘要
Developments in large language models (LLMs) in the past 2 years have shifted the focus from text, image, and audio generation to LLMs capable of multistep reasoning (thinking). The development of LLMs is particularly important for medicine and health care, but the translation of these models has been limited by the black-box nature of previous LLMs. New reasoning-driven LLMs incorporate chain-of-thought prompting and reveal intermediate reasoning steps, offering transparency and traceability, potentially improving the clinical adoption and utility of LLMs. In this Viewpoint, we examine four emerging reasoning-driven LLMs, namely OpenAI's o1 and o3-mini, Google's Gemini 2.0 Flash Thinking, and DeepSeek R1. We compare their methodological approaches, benchmark their performance on medical question-answering tasks, and assess their potential for clinical integration. We highlight both opportunities and challenges associated with deploying reasoning-driven LLMs. Key future considerations include real-world validation, rigorous benchmarking with ethical safeguards, and advancements in improving the efficiency and sustainability of reasoning-driven LLMs. Addressing these challenges will enable the fine-tuning of these LLMs for specific medical applications, enhancing their potential clinical decision support, patient education, medical training, and evidence synthesis.
科研通智能强力驱动
Strongly Powered by AbleSci AI