1 paper
Suhwan Choi, Yunsung Lee, Yubeen Park +4
Vision-Language-Action (VLA) models are increasingly evaluated across multiple simulation benchmarks, yet adding each benchmark to an evaluation pipeline requires resolving incompa…