课题基金 / 基金详情

Person-specific automatic speaker recognition: understanding the behaviour of individual speakers for applications of ASR

Person-specific automatic speaker recognition: understanding the behaviour of individual speakers for applications of ASR
特定于人的自动说话人识别:了解单个说话人的行为以用于 ASR 的应用
批准号:
ES/W001241/1
负责人:
Vincent Hughes
金额:
$103.22万
依托单位:
依托单位国家:
英国
项目类别:
Research Grant
财政年份:
2022
资助国家:
英国
项目状态:
未结题
起止时间:
2022 至 --

项目摘要

项目成果

Vincent Hughes的其他基金

相似基金

相关文献

中文摘要
翻译
点击翻译按钮获取中文摘要
英文摘要
Automatic speaker recognition (ASR) software processes and analyses speech to make decisions about whether two voices belong to the same or different individuals. Such technology is becoming an increasingly important part of our lives; used as a security measure when gaining access to personal accounts (e.g. banks), or as a means of tailoring content to a specific person on smart devices. Around the world, ASR systems are commonly used for investigative and forensic purposes, to analyse recordings of criminal voices where identity is unknown. Yet systems perform better or worse with certain voices. Therefore, a fundamental question remains: what makes a particular voice easy or difficult for ASR to recognise?State-of-the-art systems, using techniques from artificial intelligence (AI), have shown marked improvements in performance compared with older approaches. However, there remain issues. Firstly, ASR research has focused on minimising the effects of well-known technical factors, such as channel (e.g. mobile vs. landline telephone), recording quality and microphones. In resolving these technical challenges, large improvements in systems have been achieved. Yet little is known about how speakers themselves affect ASR performance. Secondly, ASR research has been interested in reducing overall error rates. Yet, in the real-world (where innocence and guilt may be at stake), the key question is: what is the chance the system has made an error in this specific instance? Finally, while AI approaches have undoubtedly brought improvements in overall performance, such algorithms make it more difficult to know what information systems are relying on to make decisions. This is problematic for forensic experts, who must explain their methods to non-expert end users, such as judges, juries, lawyers and police.This project is the first to systematically assess how individual speakers perform within and across ASR systems and to compare speaker effects, in terms of linguistic properties of voices or speaker demographics (e.g. accent, ethnicity, gender), with well-studied technical effects. The aim is to use this knowledge to improve ASR systems by flagging potentially problematic speakers and to develop methods to handle these problematic speakers. We will use novel, interdisciplinary methods, bringing together expertise from speech technology, linguistics, and forensic speech science. Our collaboration with commercial ASR vendor Oxford Wave Research allows us to adapt and change systems to assess the effects on results for individual speakers. We will also use highly controlled, small-scale experiments to assess speaker effects in isolation, as well as using much larger datasets of more forensically realistic recordings, provided by our project partners, the UK Ministry of Defence and the Netherlands Forensic Institute. The availability of a variety of datasets also allows us to assess the generalisability of results across a range of voices. This project is entirely driven by real-world issues and so the results will deliver considerable impact to a wide range of stakeholders. By understanding more about individuals, our results have the capability to improve overall ASR performance. This will be of benefit to users and developers of ASR systems. The results will also have specific implications for forensic and investigative applications, guiding data collection for validating methods (something which experts are under increasing regulatory pressure to do) and provide a framework for combining ASR and linguistic analysis. In doing so, through engagement with the legal community, we aim to affect a change in the status of ASR in England and Wales, such that it is admissible as expert evidence. We will deliver impact via knowledge exchange with a Forensic Advisory Panel consisting of representatives from forensic speech science, law enforcement, and the legal community.
期刊论文(2)
专著(0)
科研奖励(0)
会议论文
Reducing uncertainty at the score-to-LR stage in likelihood ratio-based forensic voice comparison using automatic speaker recognition systems
使用自动说话人识别系统减少基于似然比的法医语音比较中分数到 LR 阶段的不确定性
DOI: 10.21437/interspeech.2022-518
发表时间: 2022
期刊:
影响因子: --
作者: [Wang B]
通讯作者: Wang B
Humans and machines: novel methods for testing speaker recognition performance
  • 批准号:
    AH/T012978/1
  • 项目类别:
    Research Grant
  • 资助金额:
    $25.63万
  • 财政年份:
    2021
  • 负责人:
    Vincent Hughes
  • 依托单位:
国内基金
海外基金
新生儿坏死性小肠结肠炎中去泛素化酶USP15调控ILC3分化损伤肠道粘膜屏障的致病机制研究
  • 批准号:
    82371711
  • 项目类别:
    面上项目
  • 资助金额:
    49.00万元
  • 批准年份:
    2023
  • 负责人:
    吕志宝
  • 依托单位:
人巨细胞病毒编码蛋白UL23调控 HCMV-specific T 细胞增殖、活性及分化的机理
  • 批准号:
    32070149
  • 项目类别:
    面上项目
  • 资助金额:
    58.0万元
  • 批准年份:
    2020
  • 负责人:
    李弘剑
  • 依托单位:
花胶鱼类物种Species-specific PCR和Multiplex PCR鉴定体系研究
  • 批准号:
    31902373
  • 项目类别:
    青年科学基金项目
  • 资助金额:
    23.0万元
  • 批准年份:
    2019
  • 负责人:
    曾玲
  • 依托单位:
Dravet综合征基因突变分析及突变来源研究
  • 批准号:
    81171221
  • 项目类别:
    面上项目
  • 资助金额:
    58.0万元
  • 批准年份:
    2011
  • 负责人:
    张月华
  • 依托单位: