6 citations · 6 across the 2 of their papers we have counts for
Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
VANTAGE-Bench: Evaluating the Infrastructure AI Gap in Vision-Language Models
Zaid Pervaiz Bhat, Nimra Nayyar, Arihant Jain +6
As Vision-Language Models (VLMs) advance toward physical deployment, the focus has remained on action-oriented Embodied AI evaluated on subject-centric consumer video. This overloo…
cs.CV2024
Motor Focus: Fast Ego-Motion Prediction for Assistive Visual Navigation
Hao Wang, Jiayou Qin, Xiwen Chen +4
Assistive visual navigation systems for visually impaired individuals have become increasingly popular thanks to the rise of mobile computing. Most of these devices work by transla…
cs.CV2024★ 6 cited
VisionGPT: LLM-Assisted Real-Time Anomaly Detection for Safe Visual Navigation
Hao Wang, Jiayou Qin, Ashish Bastola +4
This paper explores the potential of Large Language Models(LLMs) in zero-shot anomaly detection for safe visual navigation. With the assistance of the state-of-the-art real-time op…