1 citations · 1 across the 6 of their papers we have counts for
1 paper · 1 filter
Zhifeng Wang, Qixuan Zhang, Peter Zhang +5
Vision Large Language Models (VLLMs) exhibit promising potential for multi-modal understanding, yet their application to video-based emotion recognition remains limited by insuffic…