2 citations · 2 across the 4 of their papers we have counts for
1 paper · 1 filter
Jiayi Chen, Wenxuan Song, Jiaxin Fang +12
Vision-Language-Action (VLA) models that encode actions using a discrete tokenization scheme have been widely adopted for robotic manipulation, but existing decoding paradigms rema…