1 paper
Chen Li, Zhantao Yang, Han Zhang +4
Vision-Language-Action (VLA) models show promise in embodied reasoning, yet remain far from true generalists-they often require task-specific fine-tuning, incur high compute costs,…