2 papers
cs.CL2025
EmoGist: Efficient In-Context Learning for Visual Emotion Understanding
Ronald Seoh, Dan Goldwasser
In this paper, we introduce EmoGist, a training-free, in-context learning method for performing visual emotion classification with LVLMs. The key intuition of our approach is that…
cs.CV2025
MagiC: Evaluating Multimodal Cognition Toward Grounded Visual Reasoning
Chengfei Wu, Ronald Seoh, Bingxuan Li +3
Recent advances in large vision-language models have led to impressive performance in visual question answering and multimodal reasoning. However, it remains unclear whether these…