1 paper · 1 filter
Songjun Tu, Qichao Zhang, Jingbo Sun +6
While multimodal large language models excel at tasks that integrate visual perception with symbolic reasoning, their performance is often undermined by a critical vulnerability: p…