课题基金 / 基金详情

Unifying audio signal processing and machine learning: a fundamental framework for machine hearing

Unifying audio signal processing and machine learning: a fundamental framework for machine hearing
统一音频信号处理和机器学习:机器听力的基本框架
批准号:
EP/L000776/1
负责人:
Richard Turner
金额:
$12.37万
依托单位:
依托单位国家:
英国
项目类别:
Research Grant
财政年份:
2013
资助国家:
英国
项目状态:
已结题
起止时间:
2013 至 --

项目摘要

项目成果

Richard Turner的其他基金

相似基金

相关文献

中文摘要
翻译
点击翻译按钮获取中文摘要
英文摘要
Modern technology is leading to a flood of audio data. For example, over seventy two hours of unstructured and unlabelled sound-tracks are uploaded to internet sites every minute. Automatic systems are urgently needed for recognising audio content so that these sound-tracks can be tagged for categorisation and search. Moreover, an increasing proportion of recordings are made on hand-held devices in challenging environments that contain multiple sound sources and noise. Such uncurated and noisy data necessitate automatic systems for cleaning the audio content and separating sources from mixtures. On a related note, devices for the hearing impaired currently perform poorly in noise. In fact, this is a major reason why six million people in the UK who would benefit from a hearing aid, do not use them (a market worth £18 billion p.a.). Patients fitted with cochlear implants suffer from similar limitations, and as the population ages more people are affected. It is clear that audio recognition and enhancement methods are required to stop us drowning in audio-data, for processing in hearing devices, and tosupport new technological innovations. Current approaches to these problems use a combination of audio signal processing (which places the audio data into a convenient format and reduces the data-rate) and machine learning (which removes noise, separates sources, or classifies the content). It is widely believed that these two fields must become increasingly integrated in the future. However, this union is currently a troubled one, suffering from four problems. Inefficiency: The methods are too inefficient when we have vast amounts of data (as is the case for audio-tracks on the web) or for real-time applications (such as is necessary in hearing aids)Impoverished models: The machine learning modules tend to be statistically limited.Unadapted: The signal processing modules are unadapted despite evidence from other fields, like computer vision, which suggests that automatic tuning leads to significant performance gains Distorted mixtures: The signal processing modules introduce non-linear distortions which are not captured by the machine learning modules.In this project we address these four limitations by introducing a new theoretical framework which unifies signal processing and machine learning. The key step is to view the signal processing module as solving an inference problem. Since the machine-learning modules are often framed in this way, the two modules can be integrated into a single coherent approach allowing technologies from the two fields to be completely integrated. In the project we will then use the new approach to develop efficient, rich, adaptive, and distortion free approaches to audio denoising, source separation and recognition. We will evaluate the the noise reduction and source separations algorithms on the hearing impaired, and the audio recognition algorithms on audio-sound track data.We believe this new framework will form a foundation of the emerging field of machine hearing. In the future, machine hearing will be deployed in a vast range of applications from music processing tasks to augmented reality systems (in conjunction with technologies from computer vision). We believe that this project will kick start this proliferation.
期刊论文(10)
专著(0)
科研奖励(0)
会议论文
DOI: --
发表时间: 2018-11
期刊:
影响因子: --
作者: [A. Solin;J. Hensman;Richard E. Turner]
通讯作者: A. Solin;J. Hensman;Richard E. Turner
Sparse Gaussian Process Variational Autoencoders
稀疏高斯过程变分自动编码器
DOI: 10.48550/arxiv.2010.10177
发表时间: 2020
期刊:
影响因子: --
作者: [Ashman M]
通讯作者: Ashman M
DOI: 10.17863/cam.15597
发表时间: 2015-04
期刊:
影响因子: --
作者: [A. G. Matthews;J. Hensman;Richard E. Turner;Zoubin Ghahramani]
通讯作者: A. G. Matthews;J. Hensman;Richard E. Turner;Zoubin Ghahramani
DOI: --
发表时间: 2018-09
期刊:
影响因子: --
作者: [Anqi Wu;Sebastian Nowozin;Edward Meeds;Richard E. Turner;José Miguel Hernández-Lobato;Alexander L. Gaunt]
通讯作者: Anqi Wu;Sebastian Nowozin;Edward Meeds;Richard E. Turner;José Miguel Hernández-Lobato;Alexander L. Gaunt
9
    Machine Learning for Tomorrow: Efficient, Flexible, Robust and Automated
    • 批准号:
      EP/T005637/1
    • 项目类别:
      Research Grant
    • 资助金额:
      $208.89万
    • 财政年份:
      2020
    • 负责人:
      Richard Turner
    • 依托单位:
    Nanoporous polymer particles and gels containing functionalized semi-rigid copolymer structures
    Machine Learning for Hearing Aids: Intelligent Processing and Fitting
    • 批准号:
      EP/M026957/1
    • 项目类别:
      Research Grant
    • 资助金额:
      $72.04万
    • 财政年份:
      2015
    • 负责人:
      Richard Turner
    • 依托单位:
    Sterically Congested and Stiffened Alternating Copolymers:  Synthesis, Solution and Solid-State Properties
    海外基金