1 paper
Youngwan Lee, Soojin Jang, Yoorhim Cho +3
Spatial reasoning is foundational for Vision-Language Models (VLMs), particularly when deployed as Vision-Language-Action (VLA) agents in physical environments. However, existing b…