1 paper
Hisayuki Yokomizo, Taiki Miyanishi, Yan Gang +3
Vision-Language Models (VLMs) are increasingly applied to robotic perception and manipulation, yet their ability to infer physical properties required for manipulation remains limi…