arXiv · 2203.13032
Multi-modal Emotion Estimation for in-the-wild Videos
Abstract
In this paper, we briefly introduce our submission to the Valence-Arousal Estimation Challenge of the 3rd Affective Behavior Analysis in-the-wild (ABAW) competition. Our method utilizes the multi-modal information, i.e., the visual and audio information, and employs a temporal encoder to model the temporal context in the videos. Besides, a smooth processor is applied to get more reasonable predictions, and a model ensemble strategy is used to improve the performance of our proposed method. The experiment results show that our method achieves 65.55% ccc for valence and 70.88% ccc for arousal on the validation set of the Aff-Wild2 dataset, which prove the effectiveness of our proposed method.
Explore related subjects
Keep this discovery
Liyu Meng, Yuchen Liu, Xiaolong Liu, Zhaopei Huang, Yuan Cheng, Meng Wang, Chuanhe Liu, Qin Jin. 2022-03-24. Multi-modal Emotion Estimation for in-the-wild Videos. https://arxiv.org/abs/2203.13032
Cite the original work for its findings. Save a collection to share your selection of sources.