1 citations · 1 across the 16 of their papers we have counts for
1 paper · 1 filter
Xingjian Tao, Yiwei Wang, Yujun Cai +2
Multi-view spatial reasoning remains difficult for current vision-language models. Even when multiple viewpoints are available, models often underutilize cross-view relations and i…