1 citations · 1 across the 5 of their papers we have counts for
1 paper · 1 filter
Furong Jia, Ling Dai, Wenjin Deng +4
Large Vision-Language Models (LVLMs) have demonstrated strong reasoning capabilities in geo-localization, yet they often struggle in real-world scenarios where visual cues are spar…