课题基金 / 基金详情

AI-driven Sound Analysis

AI-driven Sound Analysis
人工智能驱动的声音分析
批准号:
2436014
负责人:
金额:
$0.0万
依托单位:
依托单位国家:
英国
项目类别:
Studentship
财政年份:
2020
资助国家:
英国
项目状态:
未结题
起止时间:
2020 至 --
关键词:

项目摘要

项目成果

相似基金

相关文献

中文摘要
翻译
该项目将研究提出用于声音分析和合成的新机器学习方法。这包括从混合录音中分离不同的声源,以及重新制作和合成新的场景。根据计算学院对博士生的指导,学生将在前四个月主要关注文献综述,并在第四个月结束时制作第一份正式报告。该报告应是一个全面的文献综述对拟议的研究课题,并确定几个可能的研究方向。在第九个月结束时,学生将产生第一年的报告,服务于第一年年底的转移viva的目的。该报告将包括博士论文的正式介绍和陈述,全面的文献综述,目前的工作进展和未来的计划,博士学位的其余部分。博士学位的进展将从第二年开始每年进行审查。第一阶段的研究将集中在声音分析,特别是如何从母带/混合音频文件中分离出不同的声源。从混合音频文件中提取不同的声音将是这个阶段的主要目标。第二阶段将集中于将声音提取方法扩展到不同的应用中。这些应用可能包括从现有录音中重新录制音轨,将提取的声音合成新的音轨,为不同的应用(例如虚拟现实)生成动态声音。这项研究正好符合EPSRC的研究领域“视觉,听觉和其他感官”,跨越多个领域,包括人工智能技术和人机交互,以及音乐和声学技术。
英文摘要
The project will look into proposing new machine learning methods for sound analysis and synthesis. This includes separating different sound sources from mixed recordings, and remastering and synthesising new scenarios. In line with the School of Computing guidance for PhD students, the student will mainly focus on literature review in the first four months and produce the first formal report at the end of the fourth month. The report should be a comprehensive literature review on the proposed research topic and also identify a few possible research directions.At the end of the ninth month, the student will produce the first-year report which serves the purpose of the transfer viva by the end of the first year. The report will include the formal introduction and statement of the PhD topic, a comprehensive literature review, current work progresses and future plans for the remaining of the PhD.The PhD progress will be reviewed annually from the second year.The first stage of the research will focus on sound analysis, especially in how to isolate different sound sources from mastered/mixed audio files. Extracting different sounds from mixed audio files will be the main goal of this stage. New machine learning methods will be proposed for this purpose.The second stage will focus on extending the sound extraction methods into different applications. Such applications might include re-mastering the sound tracks from existing recordings, synthesising new tracks with extracted sounds, dynamic sound generation for different applications, e.g. Virtual Reality.The research fits squarely into EPSRC's research area 'Vision, hearing and other senses', spanning across multiple areas including Artificial Intelligence Technologies and Human-Computer Interaction, and Music and Acoustic Technology.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
国内基金
海外基金
Data-driven Recommendation System Construction of an Online Medical Platform Based on the Fusion of Information
基于Cache的远程计时攻击研究