Showing cs.AIShow all
2 papers · 1 filter
cs.AI2025
Investigating The Functional Roles of Attention Heads in Vision Language Models: Evidence for Reasoning Modules
Yanbei Jiang, Xueqi Ma, Shu Liu +5
Despite excelling on multimodal benchmarks, vision-language models (VLMs) largely remain a black box. In this paper, we propose a novel interpretability framework to systematically…
cs.AI2025
MEF: A Systematic Evaluation Framework for Text-to-Image Models
Xiaojing Dong, Weilin Huang, Liang Li +6
Rapid advances in text-to-image (T2I) generation have raised higher requirements for evaluation methodologies. Existing benchmarks center on objective capabilities and dimensions,…