1 citations · 1 across the 2 of their papers we have counts for
3 papers
cs.CV2026
Evolution of Accuracy and Visual-Cognitive Errors in a Decade of Vision-Language AI Models
Shravan Murlidaran, Miguel P. Eckstein
Vision language models (VLMs) have made remarkable progress in visual reasoning during the last decade. Most evaluations have used simple scenes (MS-COCO) that do not showcase comp…
cs.CV2026★ 1 cited
Why We Look Where We Look: Emergent Human-like Fixations of a Foveated Visual Language Model Maximizing Scene Understanding
Shravan Murlidaran, Ziqi Wen, Sana Shehabi +1
When humans view scenes without a specific task (free-viewing), they initially direct their eye movements toward the scene center and then fixate on people, text, objects being gaz…
cs.CV2025
Predicting Reaction Time to Comprehend Scenes with Foveated Scene Understanding Maps
Ziqi Wen, Jonathan Skaza, Shravan Murlidaran +2
Although models exist that predict human response times (RTs) in tasks such as target search and visual discrimination, the development of image-computable predictors for scene und…