79 citations · 106 across the 8 of their papers we have counts for
Showing cs.MMShow all
2 papers · 1 filter
cs.MM2022★ 1 cited
A Pre-trained Audio-Visual Transformer for Emotion Recognition
Minh Tran, Mohammad Soleymani
In this paper, we introduce a pretrained audio-visual Transformer trained on more than 500k utterances from nearly 4000 celebrities from the VoxCeleb2 dataset for human behavior un…
cs.MM2019★ 79 cited
Affective Computing for Large-Scale Heterogeneous Multimedia Data: A Survey
Sicheng Zhao, Shangfei Wang, Mohammad Soleymani +2
The wide popularity of digital photography and social networks has generated a rapidly growing volume of multimedia data (i.e., image, music, and video), resulting in a great deman…