1 citations · 1 across the 1 of their papers we have counts for
1 paper · 1 filter
Andy Zhai, Brae Liu, Bruno Fang +17
While foundation models show remarkable progress in language and vision, existing vision-language models (VLMs) still have limited spatial and embodiment understanding. Transferrin…