6 citations · 7 across the 5 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
How Much Does It Cost to Answer My Question? Benchmarking Cloud VLM-based VQA Systems
Henri Vanhuynegem, Weitao Xu, Yiran Shen +1
Vision-language models (VLMs) are becoming a practical backend for mobile visual question answering (VQA) systems, enabling smartphones and smart glasses to answer users' questions…
cs.CV2022★ 1 cited
FreeGaze: Resource-efficient Gaze Estimation via Frequency Domain Contrastive Learning
Lingyu Du, Guohao Lan
Gaze estimation is of great importance to many scientific fields and daily applications, ranging from fundamental research in cognitive psychology to attention-aware mobile systems…