1 paper
Alexander Hackett, Arnaud Denis-Remillard, Axel Cassou
How much of a vision-language model's (VLM) spatial understanding remains after the action post-training process of building a vision-language-action model (VLA)? We probe depth pe…