1 paper
Yihang Du, Juhao Liang, Zhengzhao Lai +2
Multimodal large language models (MLLMs) exhibit substantial performance degradation in non-English visual reasoning, despite the strong multilingual competence of their text-only…