Noise-Robust Key-Phrase Detectors for Automated Classroom Feedback
Noise-Robust Key-Phrase Detectors for Automated Classroom Feedback
复制标题
DOI:
10.1109/icassp40776.2020.9053173
复制
发表时间:
2020-05
期刊:
影响因子:
--
通讯作者:
Brian Zylich;J. Whitehill
中科院分区:
文献类型:
--
作者:
Brian Zylich;J. Whitehill
With the goal of giving teachers automated feedback about their classrooms, we investigate how to train automatic speech detectors of key phrases such as good job, thank you, please, and you’re welcome. This kind of language conveys support and respect from teacher to student and is one of the behavioral markers used in the established CLASS [1] classroom observation protocol. School classrooms are noisy and contain overlapping speech, presenting a highly challenging environment for automatic speech recognition (ASR), even for state-of-the-art approaches. We train deep neural networks using hierarchical multitask learning (MTL) on a modest-sized but highly-tailored dataset of classroom speech. Compared to 2 state-of-the-art ASR systems for general-purpose speech recognition (Google [2] and Deep-Speech [3]), our system delivers a substantially improved recall rate (50.4% versus 20.5%) while matching their precision (30%). Moreover, our system’s predictions correlate with several dimensions of the CLASS.