中文

面向 SE&R 2022 挑战赛的领域特定 Wav2vec 2.0 微调

计算与语言 2022-08-01 v1 声音 音频与语音处理

摘要

本文介绍了我们为共享任务“葡萄牙语自发与预备语音自动语音识别及语音情感识别(SE&R 2022)”构建鲁棒 ASR 模型的努力。该挑战赛旨在推进葡萄牙语的 ASR 研究,考虑不同方言中的预备语音与自发语音。我们的方法采用领域特定的方式微调 ASR 模型,应用增益归一化与选择性噪声插入。所提方法在可用 4 个赛道中的 3 个上优于提供的强基线测试集结果。

关键词

引用

@article{arxiv.2207.14418,
  title  = {Domain Specific Wav2vec 2.0 Fine-tuning For The SE&R 2022 Challenge},
  author = {Alef Iury Siqueira Ferreira and Gustavo dos Reis Oliveira},
  journal= {arXiv preprint arXiv:2207.14418},
  year   = {2022}
}

备注

Proceedings of the First Workshop on Automatic Speech Recognition for Spontaneous and Prepared Speech & Speech Emotion Recognition in Portuguese (SE&R 2022), co-located with PROPOR 2022