Natural Speech Technology
Natural Speech Technology
批准号:
EP/I031022/1
负责人:
Steve Renals
金额:
$794.6万
依托单位:
依托单位国家:
英国
项目类别:
Research Grant
财政年份:
2011
资助国家:
英国
项目状态:
已结题
起止时间:
2011 至 --
中文摘要
点击翻译按钮获取中文摘要
英文摘要
Humans are highly adaptable, and speech is our natural medium for informal communication. When communicating, we continuously adjust to other people, to the situation, and to the environment, using previously acquired knowledge to make this adaptation seem almost instantaneous. Humans generalise, enabling efficient communication in unfamiliar situations and rapid adaptation to new speakers or listeners. Current speech technology works well for certain controlled tasks and domains, but is far from natural, a consequence of its limited ability to acquire knowledge about people or situations, to adapt, and to generalise. This accounts for the uneasy public reaction to speech-driven systems. For example, text-to-speech synthesis can be as intelligible as human speech, but lacks expression and is not perceived as natural. Similarly, the accuracy of speech recognition systems can collapse if the acoustic environment or task domain changes, conditions which a human listener would handle easily. Research approaches to these problems have hitherto been piecemeal and as a result progress has been patchy. In contrast NST will focus on the integrated theoretical development of new joint models for speech recognition and synthesis. These models will allow us to incorporate knowledge about the speakers, the environment, the communication context and awareness of the task, and will learn and adapt from real world data in an online, unsupervised manner. This theoretical unification is already underway within the NST labs and, combined with our record of turning theory into practical state-of-the-art applications, will enable us to bring a naturalness to speech technology that is not currently attainable.The NST programme will yield technology which (1) approaches human adaptability to new communication situations, (2) is capable of personalised communication, and (3) takes account of speaker intention and expressiveness in speech recognition and synthesis. This is an ambitious vision. Its success will be measured in terms of how the theoretical development reshapes the field over the next decade, the takeup of the software systems that we shall develop, and through the impact of our exemplar interactive applications.We shall establish a strong User Group to maximise the impact of the project, with a members concerned with clinical applications, as well as more general speech technology. Members of the User Group include Toshiba, EADS Innovation Works, Cisco, Barnsley Hospital NHS Foundation Trust, and the Euan MacDonald Centre for MND Research. An important interaction with the User Group will be validating our systems on their data and tasks, discussed at an annual user workshop.
期刊论文(10)
专著(0)
科研奖励(0)
会议论文
登录
查看更多内容
Multi-reference WER for evaluating ASR for languages with no orthographic rule
用于评估没有拼写规则的语言的 ASR 的多参考 WER
DOI:
--
发表时间:
2015
期刊:
Proc IEEE ASRU
影响因子:
--
作者:
[Ali A]
通讯作者:
Ali A
Reactive accent interpolation through an interactive map application
通过交互式地图应用程序进行反应式重音插值
DOI:
--
发表时间:
2013
期刊:
Proceedings of the Annual Conference of the International Speech Communication Association, INTERSPEECH
影响因子:
--
作者:
[Astrinaki M.]
通讯作者:
Astrinaki M.
A system for automatic alignment of broadcast media captions using weighted finite-state transducers
DOI:
10.1109/asru.2015.7404861
发表时间:
2015-12
期刊:
2015 IEEE Workshop on Automatic Speech Recognition and Understanding (ASRU)
影响因子:
--
作者:
[P. Bell;S. Renals]
通讯作者:
P. Bell;S. Renals
DOI:
10.1016/j.specom.2011.08.001
发表时间:
2012-02
期刊:
Speech Commun.
影响因子:
--
作者:
[Sebastian Andersson;J. Yamagishi;R. Clark]
通讯作者:
Sebastian Andersson;J. Yamagishi;R. Clark
DOI:
10.21437/interspeech.2014-320
发表时间:
2014
期刊:
影响因子:
--
作者:
[M. Aylett;R. Dall;Arnab Ghoshal;G. Henter;Thomas Merritt]
通讯作者:
M. Aylett;R. Dall;Arnab Ghoshal;G. Henter;Thomas Merritt
共 8 条
SpeechWave
-
批准号:EP/R012180/1
-
项目类别:Research Grant
-
资助金额:$85.12万
-
财政年份:2018
-
负责人:Steve Renals
-
依托单位:
Ultrax2020: Ultrasound Technology for Optimising the Treatment of Speech Disorders.
-
批准号:EP/P02338X/1
-
项目类别:Research Grant
-
资助金额:$122.92万
-
财政年份:2017
-
负责人:Steve Renals
-
依托单位:
Ultrax: Real-time tongue tracking for speech therapy using ultrasound
-
批准号:EP/I027696/1
-
项目类别:Research Grant
-
资助金额:$74.69万
-
财政年份:2011
-
负责人:Steve Renals
-
依托单位:
MultiMemoHome: Multimodal Reminders Within the Home
-
批准号:EP/G060614/1
-
项目类别:Research Grant
-
资助金额:$31.35万
-
财政年份:2009
-
负责人:Steve Renals
-
依托单位:
Data-driven articulatory modelling: foundations for a new generation of speech synthesis
-
批准号:EP/E027741/1
-
项目类别:Research Grant
-
资助金额:$36.55万
-
财政年份:2006
-
负责人:Steve Renals
-
依托单位:
海外基金