AI chatbots not yet ready for clinical use.

AI chatbots not yet ready for clinical use.
复制标题

DOI:
10.3389/fdgth.2023.1161098
复制
发表时间:
2023
影响因子:
--
通讯作者:
Teo, James T. T.
Teo, James T. T.
中科院分区:
其他
文献类型:
--
作者:
Au Yeung, Joshua;Kraljevic, Zeljko;Luintel, Akish;Balston, Alfred;Idowu, Esther;Dobson, Richard J. J.;Teo, James T. T.

文献摘要

参考文献

被引文献

相似文献

随着大型语言模型(llm)的扩展和变得更加先进,会话人工智能(或“聊天机器人”)的自然语言处理能力也在不断增强。OpenAI最近发布的ChatGPT使用基于转换器的模型,在一般领域知识上实现类似人类的文本生成和问答,而GatorTron等医疗保健特定大型语言模型(LLM)则专注于现实世界的医疗保健领域知识。随着法学硕士在医学问题和回答基准上取得接近人类水平的表现,会话人工智能很可能很快就会被开发出来用于医疗保健领域。在本文中,我们讨论了两种不同的生成预训练转换器方法的潜力并比较了它们的性能——chatgpt,最广泛使用的一般会话LLM,以及Foresight,一种基于GPT(生成预训练转换器)的模型,专注于对患者和疾病进行建模。对基于临床影像预测相关诊断的任务进行了比较。我们还讨论了基于变压器的聊天机器人在临床应用中的重要考虑因素和局限性。
As large language models (LLMs) expand and become more advanced, so do the natural language processing capabilities of conversational AI, or “chatbots”. OpenAI's recent release, ChatGPT, uses a transformer-based model to enable human-like text generation and question-answering on general domain knowledge, while a healthcare-specific Large Language Model (LLM) such as GatorTron has focused on the real-world healthcare domain knowledge. As LLMs advance to achieve near human-level performances on medical question and answering benchmarks, it is probable that Conversational AI will soon be developed for use in healthcare. In this article we discuss the potential and compare the performance of two different approaches to generative pretrained transformers—ChatGPT, the most widely used general conversational LLM, and Foresight, a GPT (generative pretrained transformer) based model focused on modelling patients and disorders. The comparison is conducted on the task of forecasting relevant diagnoses based on clinical vignettes. We also discuss important considerations and limitations of transformer-based chatbots for clinical use.
DOI: 10.1016/j.chb.2011.09.006
发表时间: 2012-01-01
影响因子: 9.9
作者:
Kim, Youjeong;Sundar, S. Shyam
通讯作者: Sundar, S. Shyam
DOI: 10.1038/s41598-020-62922-y
发表时间: 2020-04-28
期刊: SCIENTIFIC REPORTS
影响因子: 4.6
作者:
Li, Yikuan;Rao, Shishir;Salimi-Khorshidi, Gholamreza
通讯作者: Salimi-Khorshidi, Gholamreza
DOI: 10.1371/journal.pone.0159224
发表时间: 2016-08-08
期刊: PLOS ONE
影响因子: 3.7
作者:
Singhal, Astha;Tien, Yu-Yu;Hsia, Renee Y.
通讯作者: Hsia, Renee Y.
DOI: 10.1126/science.aal4230
发表时间: 2017-04-14
期刊: SCIENCE
影响因子: 56.9
作者:
Caliskan, Aylin;Bryson, Joanna J.;Narayanan, Arvind
通讯作者: Narayanan, Arvind
为了分类和诊断目的,人工智能和人类医生的比较。
DOI: 10.3389/frai.2020.543405
发表时间: 2020
影响因子: 4
作者:
Baker A;Perov Y;Middleton K;Baxter J;Mullarkey D;Sangar D;Butt M;DoRosario A;Johri S
通讯作者: Johri S