9 citations · 11 across the 3 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2023★ 1 cited
Enhancing image captioning with depth information using a Transformer-based framework
Aya Mahmoud Ahmed, Mohamed Yousef, Khaled F. Hussain +1
Captioning images is a challenging scene-understanding task that connects computer vision and natural language processing. While image captioning models have been successful in pro…
cs.CV2018★ 9 cited
Accurate, Data-Efficient, Unconstrained Text Recognition with Convolutional Neural Networks
Mohamed Yousef, Khaled F. Hussain, Usama S. Mohammed
Unconstrained text recognition is an important computer vision task, featuring a wide variety of different sub-tasks, each with its own set of challenges. One of the biggest promis…