2 papers
cs.AI2025
Investigating The Functional Roles of Attention Heads in Vision Language Models: Evidence for Reasoning Modules
Yanbei Jiang, Xueqi Ma, Shu Liu +5
Despite excelling on multimodal benchmarks, vision-language models (VLMs) largely remain a black box. In this paper, we propose a novel interpretability framework to systematically…
cs.CV2025
Reasoning Like Experts: Leveraging Multimodal Large Language Models for Drawing-based Psychoanalysis
Xueqi Ma, Yanbei Jiang, Sarah Erfani +4
Multimodal Large Language Models (MLLMs) have demonstrated exceptional performance across various objective multimodal perception tasks, yet their application to subjective, emotio…