1 paper
Taishan Li, Jiwen Zhang, Siyuan Wang +2
Vision-Language-Action (VLA) models achieve strong performance on standard manipulation benchmarks, but most evaluations assume that task-relevant objects are fully visible. This a…