8 citations · 8 across the 3 of their papers we have counts for
3 papers
Learning Saliency From Fixations
Yasser Abdelaziz Dahou Djilali, Kevin McGuiness, Noel O'Connor
We present a novel approach for saliency prediction in images, leveraging parallel decoding in transformers to learn saliency solely from fixation maps. Models typically rely on co…
Do VSR Models Generalize Beyond LRS3?
Yasser Abdelaziz Dahou Djilali, Sanath Narayan, Eustache Le Bihan +3
The Lip Reading Sentences-3 (LRS3) benchmark has primarily been the focus of intense research in visual speech recognition (VSR) during the last few years. As a result, there is an…
ATSal: An Attention Based Architecture for Saliency Prediction in 360 Videos
Yasser Dahou, Marouane Tliba, Kevin McGuinness +1
The spherical domain representation of 360 video/image presents many challenges related to the storage, processing, transmission and rendering of omnidirectional videos (ODV). Mode…