基于时间有序深度音频与随机视觉特征的双模态第一印象(大五人格特质)识别
计算机视觉与模式识别
2016-11-01 v1
摘要
我们提出了一种从短视频中基于大五人格特质进行第一印象识别的新方法。大五人格特质是一种利用五个宽泛类别描述人类人格的模型:外向性、宜人性、尽责性、神经质和开放性。我们训练了两种双模态端到端深度神经网络架构,使用时间有序的音频和来自少量帧的新颖随机视觉特征,且未发生过拟合。我们通过经验表明,即使仅使用输入的一小部分子集进行训练,训练好的模型也表现极佳。我们的方法在 ChaLearn LAP 2016 表观人格分析(APA)竞赛中使用 ChaLearn LAP APA2016 数据集进行了评估,并取得了优异性能。
引用
@article{arxiv.1610.10048,
title = {Bi-modal First Impressions Recognition using Temporally Ordered Deep Audio and Stochastic Visual Features},
author = {Arulkumar Subramaniam and Vismay Patel and Ashish Mishra and Prashanth Balasubramanian and Anurag Mittal},
journal= {arXiv preprint arXiv:1610.10048},
year = {2016}
}
备注
to be published in: ECCV 2016 Workshops proceedings (Apparent Personality Analysis)