1 paper
Juntao Gao, Feiyang Ye, Jing Zhang +1
Vision-Language-Action (VLA) models have emerged as a powerful paradigm in Embodied AI. However, the significant computational overhead of processing redundant visual tokens remain…