GPT-4V(ision) Unsuitable for Clinical Care and Education: A Clinician-Evaluated Assessment
GPT-4V(ision) Unsuitable for Clinical Care and Education: A Clinician-Evaluated Assessment
复制标题
GPT-4V(ision) 不适合临床护理和教育:临床医生评估
DOI:
10.1101/2023.11.15.23298575
复制
发表时间:
2023
期刊:
影响因子:
--
通讯作者:
Bo Wang
中科院分区:
文献类型:
--
作者:
Senthujan Senkaiahliyan;Augustin Toma;Jun Ma;An;Andrew Ha;Kevin R. An;Hrishikesh Suresh;Barry Rubin;Bo Wang
OpenAI's large multimodal model, GPT-4V(ision), was recently developed for general image interpretation. However, less is known about its capabilities with medical image interpretation and diagnosis. Board-certified physicians and senior residents assessed GPT-4V's proficiency across a range of medical conditions using imaging modalities such as CT scans, MRIs, ECGs, and clinical photographs. Although GPT-4V is able to identify and explain medical images, its diagnostic accuracy and clinical decision-making abilities are poor, posing risks to patient safety. Despite the potential that large language models may have in enhancing medical education and delivery, the current limitations of GPT-4V in interpreting medical images reinforces the importance of appropriate caution when using it for clinical decision-making.