1 paper · 1 filter
Zhao Yang, Yinan Shi, Mingyuan Yao +3
Vision-language action (VLA) models increasingly adopt chunked action heads to satisfy real-time constraints; however, this introduces boundary jitter: overlapping regions between…