English

Speech-to-speech Translation between Untranscribed Unknown Languages

Computation and Language 2019-10-08 v2 Machine Learning Sound Audio and Speech Processing

Abstract

In this paper, we explore a method for training speech-to-speech translation tasks without any transcription or linguistic supervision. Our proposed method consists of two steps: First, we train and generate discrete representation with unsupervised term discovery with a discrete quantized autoencoder. Second, we train a sequence-to-sequence model that directly maps the source language speech to the target language's discrete representation. Our proposed method can directly generate target speech without any auxiliary or pre-training steps with a source or target transcription. To the best of our knowledge, this is the first work that performed pure speech-to-speech translation between untranscribed unknown languages.

Keywords

Cite

@article{arxiv.1910.00795,
  title  = {Speech-to-speech Translation between Untranscribed Unknown Languages},
  author = {Andros Tjandra and Sakriani Sakti and Satoshi Nakamura},
  journal= {arXiv preprint arXiv:1910.00795},
  year   = {2019}
}

Comments

Accepted in IEEE ASRU 2019. Web-page for more samples & details: https://sp2code-translation-v1.netlify.com/

R2 v1 2026-06-23T11:32:25.875Z