10 citations · 14 across the 5 of their papers we have counts for
1 paper · 2 filters
Ken Deng, Yifu Qiu, Yoni Kasten +2
We study whether vision-language models (VLMs) can solve relative camera pose estimation (RCPE) from image pairs, a direct test of multi-view spatial reasoning. We cast RCPE as a d…