1 citations · 1 across the 6 of their papers we have counts for
1 paper · 1 filter
Zhanpeng Luo, Ce Zhang, Silong Yong +6
Multi-modal Large Language Models (MLLMs) have demonstrated strong capabilities in general-purpose perception and reasoning, but they still struggle with tasks that require spatial…