1 paper · 1 filter
Qitong Wang, Fan Du, Pranav Maneriker +2
The rapid rise of Vision-Language Models (VLMs) in egocentric visual understanding has made low-latency inference in human-robot collaborative (HRC) tasks increasingly critical. We…