A multi-device dataset for urban acoustic scene classification

A multi-device dataset for urban acoustic scene classification
复制标题

DOI:
--
复制
发表时间:
2018-07
期刊:
--
影响因子:
--
通讯作者:
A. Mesaros;Toni Heittola;T. Virtanen
A. Mesaros;Toni Heittola;T. Virtanen
中科院分区:
其他
文献类型:
--
作者:
A. Mesaros;Toni Heittola;T. Virtanen

文献摘要

被引文献

相似文献

本文介绍了DCASE 2018挑战赛的声学场景分类任务以及为该任务提供的TUT Urban Acoustic Scenes 2018数据集,并评估了任务中基线系统的性能。与前几年的挑战一样,该任务被定义为使用监督的闭集分类设置将短音频样本分类为预定义的声学场景类之一。新记录的TUT Urban Acoustic Scenes 2018数据集由十个不同的声学场景组成,并在六个欧洲大城市进行了记录,因此它具有比以前用于此任务的数据集更高的声学可变性,除了高质量的双耳录音外,它还包括使用移动的设备记录的数据。我们还介绍了由卷积神经网络组成的基线系统,以及使用推荐的交叉验证设置在子任务中的性能。
This paper introduces the acoustic scene classification task of DCASE 2018 Challenge and the TUT Urban Acoustic Scenes 2018 dataset provided for the task, and evaluates the performance of a baseline system in the task. As in previous years of the challenge, the task is defined for classification of short audio samples into one of predefined acoustic scene classes, using a supervised, closed-set classification setup. The newly recorded TUT Urban Acoustic Scenes 2018 dataset consists of ten different acoustic scenes and was recorded in six large European cities, therefore it has a higher acoustic variability than the previous datasets used for this task, and in addition to high-quality binaural recordings, it also includes data recorded with mobile devices. We also present the baseline system consisting of a convolutional neural network and its performance in the subtasks using the recommended cross-validation setup.