1 citations · 1 across the 2 of their papers we have counts for
1 paper · 1 filter
Jingkai Wang, Zihan Tang, Gu Zhang +7
Vision-language-action policies rely on large multimodal backbones to jointly perform perception, language conditioning, and action generation at every control step. Much of this c…