Content4All Open Research Sign Language Translation Datasets

Content4All Open Research Sign Language Translation Datasets
复制标题

DOI:
10.1109/fg52635.2021.9667087
复制
发表时间:
2021-05
期刊:
2021 16th IEEE International Conference on Automatic Face and Gesture Recognition (FG 2021)
影响因子:
--
通讯作者:
N. C. Camgoz;Ben Saunders;Guillaume Rochette;Marco Giovanelli;Giacomo Inches;Robin Nachtrab-Ribback;R. Bowden
N. C. Camgoz;Ben Saunders;Guillaume Rochette;Marco Giovanelli;Giacomo Inches;Robin Nachtrab-Ribback;R. Bowden
中科院分区:
其他
文献类型:
--
作者:
N. C. Camgoz;Ben Saunders;Guillaume Rochette;Marco Giovanelli;Giacomo Inches;Robin Nachtrab-Ribback;R. Bowden

文献摘要

被引文献

相似文献

计算手语研究缺乏能够创建有用的现实应用程序的大规模数据集。迄今为止,大多数研究仅限于小领域的原型系统,例如天气预报。为了解决这个问题并推动该领域向前发展,我们发布了六个数据集,其中包含更广泛的新闻领域的 190 小时的镜头。此后,聋人专家和口译员对 20 小时的录像进行了注释,并公开用于研究目的。在本文中,我们分享了数据集收集过程和开发的工具,以实现手语视频和字幕的对齐,以及基线翻译结果,以支持未来的研究。
Computational sign language research lacks the large-scale datasets that enables the creation of useful real-life applications. To date, most research has been limited to prototype systems on small domains of discourse, e.g. weather forecasts. To address this issue and to push the field forward, we release six datasets comprised of 190 hours of footage on the larger domain of news. From this, 20 hours of footage have been annotated by Deaf experts and interpreters and is made publicly available for research purposes. In this paper, we share the dataset collection process and tools developed to enable the alignment of sign language video and subtitles, as well as baseline translation results to underpin future research.