1 citations · 1 across the 5 of their papers we have counts for
1 paper · 1 filter
Mahtab Bigverdi, Linjie Li, Weikai Huang +8
Vision language models (VLMs) excel at many tasks but still struggle with spatial reasoning when critical information is not directly observable. Many such problems require imagina…