Creation of a Doctor-Patient Dialogue Corpus Using Standardized Patients
Creation of a Doctor-Patient Dialogue Corpus Using Standardized Patients
复制标题
使用标准化患者创建医患对话语料库
DOI:
--
复制
发表时间:
2004
期刊:
影响因子:
--
通讯作者:
S. Ganjavi
中科院分区:
文献类型:
--
作者:
Robert S. Melvin;Win May;Shrikanth S. Narayanan;P. Georgiou;S. Ganjavi
In this paper we describe the development of a doctor-patient dialogue corpus to support a speech-to-speech machine translation effort for English-Persian medical dialogues. The corpus was developed by recording and transcribing English-to-English dialogues between medical students and standardized patients (actors who have been trained to portray illness or injury victims), and then translated into Persian. We discuss some of the benefits and drawbacks to creating a corpus in this way. Benefits include the ability to customize the corpus in a way that would be infeasible for actual doctor-patient data and avoidance of privacy and legal issues, while drawbacks include the fact that the Persian does not originate as speech, but as text translation of English speech. We address concerns such as the authenticity of the dialogues and the value of such data for system development.