中文

面向野外视频的多模态情绪估计

计算机视觉与模式识别 2022-04-01 v4 图像与视频处理

摘要

本文简要介绍我们提交给第三届野外情感行为分析(ABAW)竞赛效价-唤醒度估计挑战赛的方法。我们的方法利用多模态信息,即视觉和音频信息,并采用时间编码器对视频中的时间上下文建模。此外,应用平滑处理器以获得更合理的预测,并使用模型集成策略来提升所提方法的性能。实验结果表明,我们的方法在Aff-Wild2数据集验证集上效价达到65.55% ccc、唤醒度达到70.88% ccc,证明了所提方法的有效性。

关键词

引用

@article{arxiv.2203.13032,
  title  = {Multi-modal Emotion Estimation for in-the-wild Videos},
  author = {Liyu Meng and Yuchen Liu and Xiaolong Liu and Zhaopei Huang and Yuan Cheng and Meng Wang and Chuanhe Liu and Qin Jin},
  journal= {arXiv preprint arXiv:2203.13032},
  year   = {2022}
}