Speech-to-Speech Translation Between Untranscribed Unknown Languages
Speech-to-Speech Translation Between Untranscribed Unknown Languages
复制标题
DOI:
10.1109/asru46091.2019.9003853
复制
发表时间:
2019-10
期刊:
影响因子:
--
通讯作者:
Andros Tjandra;S. Sakti;Satoshi Nakamura
中科院分区:
文献类型:
--
作者:
Andros Tjandra;S. Sakti;Satoshi Nakamura
In this paper, we explore a method for training speech-to-speech translation tasks without any transcription or linguistic supervision. Our proposed method consists of two steps: First, we train and generate discrete representation with unsupervised term discovery with a discrete quantized autoencoder. Second, we train a sequence-to-sequence model that directly maps the source language speech to the target languages discrete representation. Our proposed method can directly generate target speech without any auxiliary or pre-training steps with a source or target transcription. To the best of our knowledge, this is the first work that performed pure speech-to-speech translation between untranscribed unknown languages.