1 paper · 1 filter
Zhimin Li, Pan Wang, Jingxian Chen +4
Wearable VLM pipelines promise continuous multimodal assistance from egocentric visual capture: a user asks a task-driven question about the surrounding scene, and the system uses…