2 papers
cs.CV2025
Visual and Textual Prompts in VLLMs for Enhancing Emotion Recognition
Zhifeng Wang, Qixuan Zhang, Peter Zhang +5
Vision Large Language Models (VLLMs) exhibit promising potential for multi-modal understanding, yet their application to video-based emotion recognition remains limited by insuffic…
cs.CV2024
Visual Prompting in LLMs for Enhancing Emotion Recognition
Qixuan Zhang, Zhifeng Wang, Dylan Zhang +5
Vision Large Language Models (VLLMs) are transforming the intersection of computer vision and natural language processing. Nonetheless, the potential of using visual prompts for em…