17 citations · 59 across the 61 of their papers we have counts for
1 paper · 1 filter
Yufeng Ji, Wenhao Tang, Haoyi Niu +3
Action supervision in vision-language-action (VLA) models is often treated as a downstream objective for learning action prediction. In this paper, we study it instead as a force t…