1 paper
Zhuoran Song, Haozhe Jiang, Chunyu Qi +4
Vision-Language-Action (VLA) models have demonstrated strong potential for embodied AI, yet their high inference latency on GPUs limits real-time deployment. Existing accelerators,…