English

Assessing Evaluation Metrics for Speech-to-Speech Translation

Computation and Language 2021-10-27 v1 Sound Audio and Speech Processing

Abstract

Speech-to-speech translation combines machine translation with speech synthesis, introducing evaluation challenges not present in either task alone. How to automatically evaluate speech-to-speech translation is an open question which has not previously been explored. Translating to speech rather than to text is often motivated by unwritten languages or languages without standardized orthographies. However, we show that the previously used automatic metric for this task is best equipped for standardized high-resource languages only. In this work, we first evaluate current metrics for speech-to-speech translation, and second assess how translation to dialectal variants rather than to standardized languages impacts various evaluation methods.

Keywords

Cite

@article{arxiv.2110.13877,
  title  = {Assessing Evaluation Metrics for Speech-to-Speech Translation},
  author = {Elizabeth Salesky and Julian Mäder and Severin Klinger},
  journal= {arXiv preprint arXiv:2110.13877},
  year   = {2021}
}

Comments

ASRU 2021

R2 v1 2026-06-24T07:12:30.693Z