Interactive Visualization of AI-based Speech Recognition Texts

Interactive Visualization of AI-based Speech Recognition Texts
复制标题

DOI:
10.2312/eurova.20201091
复制
发表时间:
2020
期刊:
--
影响因子:
--
通讯作者:
Tsung Heng Wu;Ye Zhao;Md Amiruzzaman
Tsung Heng Wu;Ye Zhao;Md Amiruzzaman
中科院分区:
其他
文献类型:
--
作者:
Tsung Heng Wu;Ye Zhao;Md Amiruzzaman

文献摘要

被引文献

相似文献

语音识别技术近年来随着深度学习网络的人工智能技术而取得了令人印象深刻的成功。语音到文本工具在许多社交应用程序中变得普遍,例如field调查。然而,语音转录结果远未达到完美,无法直接供领域科学家和从业者在这些应用中使用,这阻碍了用户充分利用人工智能工具。本文通过具体的任务描述和实例说明了交互可视化在人工智能后理解、编辑和分析语音识别结果中的重要作用。
Speech recognition technology has achieved impressive success recently with AI techniques of deep learning networks. Speech-to-text tools are becoming prevalent in many social applications such as field surveys. However, the speech transcription results are far from perfection for direct use in these applications by domain scientists and practitioners, which prevents the users from fully leveraging the AI tools. In this paper, we show interactive visualization can play important roles in post-AI understanding, editing, and analysis of speech recognition results by presenting specified task characterization and case examples.