Showing cs.CVShow all
3 papers · 1 filter
cs.CV2025
GazeVLM: A Vision-Language Model for Multi-Task Gaze Understanding
Athul M. Mathew, Haithem Hermassi, Thariq Khalid +1
Gaze understanding unifies the detection of people, their gaze targets, and objects of interest into a single framework, offering critical insight into visual attention and intent…
cs.CV2025
Leveraging Multi-Modal Saliency and Fusion for Gaze Target Detection
Athul M. Mathew, Arshad Ali Khan, Thariq Khalid +2
Gaze target detection (GTD) is the task of predicting where a person in an image is looking. This is a challenging task, as it requires the ability to understand the relationship b…
cs.CV2022
Ego Vehicle Speed Estimation using 3D Convolution with Masked Attention
Athul M. Mathew, Thariq Khalid
Speed estimation of an ego vehicle is crucial to enable autonomous driving and advanced driver assistance technologies. Due to functional and legacy issues, conventional methods de…