26 citations · 30 across the 4 of their papers we have counts for
4 papers
Leveraging Visemes for Better Visual Speech Representation and Lip Reading
Javad Peymanfard, Vahid Saeedi, Mohammad Reza Mohammadi +2
Lip reading is a challenging task that has many potential applications in speech recognition, human-computer interaction, and security systems. However, existing lip reading system…
Word-level Persian Lipreading Dataset
Javad Peymanfard, Ali Lashini, Samin Heydarian +2
Lip-reading has made impressive progress in recent years, driven by advances in deep learning. Nonetheless, the prerequisite such advances is a suitable dataset. This paper provide…
A Multi-Purpose Audio-Visual Corpus for Multi-Modal Persian Speech Recognition: the Arman-AV Dataset
Javad Peymanfard, Samin Heydarian, Ali Lashini +3
In recent years, significant progress has been made in automatic lip reading. But these methods require large-scale datasets that do not exist for many low-resource languages. In t…
Learning to predict where to look in interactive environments using deep recurrent q-learning
Sajad Mousavi, Michael Schukat, Enda Howley +2
Bottom-Up (BU) saliency models do not perform well in complex interactive environments where humans are actively engaged in tasks (e.g., sandwich making and playing the video games…