面向 SE&R 2022 挑战赛的领域特定 Wav2vec 2.0 微调
计算与语言
2022-08-01 v1 声音
音频与语音处理
摘要
本文介绍了我们为共享任务“葡萄牙语自发与预备语音自动语音识别及语音情感识别(SE&R 2022)”构建鲁棒 ASR 模型的努力。该挑战赛旨在推进葡萄牙语的 ASR 研究,考虑不同方言中的预备语音与自发语音。我们的方法采用领域特定的方式微调 ASR 模型,应用增益归一化与选择性噪声插入。所提方法在可用 4 个赛道中的 3 个上优于提供的强基线测试集结果。
引用
@article{arxiv.2207.14418,
title = {Domain Specific Wav2vec 2.0 Fine-tuning For The SE&R 2022 Challenge},
author = {Alef Iury Siqueira Ferreira and Gustavo dos Reis Oliveira},
journal= {arXiv preprint arXiv:2207.14418},
year = {2022}
}
备注
Proceedings of the First Workshop on Automatic Speech Recognition for Spontaneous and Prepared Speech & Speech Emotion Recognition in Portuguese (SE&R 2022), co-located with PROPOR 2022