What Would it Take to get Biomedical QA Systems into Practice?

What Would it Take to get Biomedical QA Systems into Practice?
复制标题

DOI:
10.18653/v1/2021.mrqa-1.3
复制
发表时间:
2021-11
期刊:
Proceedings of the Conference on Empirical Methods in Natural Language Processing. Conference on Empirical Methods in Natural Language Processing
影响因子:
--
通讯作者:
--
中科院分区:
其他
文献类型:
--
作者:

文献摘要

相似文献

医疗问答(QA)系统有可能根据最新的证据,按需回答临床医生关于治疗和诊断的不确定性。然而,尽管NLP社区在一般QA方面取得了重大进展,但医学QA系统仍然没有广泛应用于临床环境。其中一个可能的原因是临床医生可能不容易信任QA系统输出,部分原因是透明度,可信度和出处在设计此类模型时并不是关键考虑因素。在本文中,我们讨论了一组标准,如果满足,我们认为可能会增加生物医学QA系统的实用性,这可能反过来导致在实践中采用这种系统。我们评估现有的模型,任务和数据集,这些标准,突出了以前提出的方法的缺点,并指出什么可能是更有用的QA系统。
Medical question answering (QA) systems have the potential to answer clinicians’ uncertainties about treatment and diagnosis on-demand, informed by the latest evidence. However, despite the significant progress in general QA made by the NLP community, medical QA systems are still not widely used in clinical environments. One likely reason for this is that clinicians may not readily trust QA system outputs, in part because transparency, trustworthiness, and provenance have not been key considerations in the design of such models. In this paper we discuss a set of criteria that, if met, we argue would likely increase the utility of biomedical QA systems, which may in turn lead to adoption of such systems in practice. We assess existing models, tasks, and datasets with respect to these criteria, highlighting shortcomings of previously proposed approaches and pointing toward what might be more usable QA systems.