1 paper
Seon Gyeom Kim, Jae Young Choi, Ryan Rossi +2
The field of Multimodal Large Language Models (MLLMs) has made remarkable progress in visual understanding tasks, presenting a vast opportunity to predict the perceptual and emotio…