1 paper · 1 filter
Jiahui Niu, Kefan Gu, Yucheng Zhao +5
Diffusion-based vision-language-action models (dVLAs) are promising for embodied intelligence but are fundamentally limited in real-time deployment by the high latency of full infe…