2 citations · 4 across the 3 of their papers we have counts for
3 papers
EgMM-Corpus: A Multimodal Vision-Language Dataset for Egyptian Culture
Mohamed Gamil, Abdelrahman Elsayed, Abdelrahman Lila +4
Despite recent advances in AI, multimodal culturally diverse datasets are still limited, particularly for regions in the Middle East and Africa. In this paper, we introduce EgMM-Co…
Invizo: Arabic Handwritten Document Optical Character Recognition Solution
Alhossien Waly, Bassant Tarek, Ali Feteha +4
Converting images of Arabic text into plain text is a widely researched topic in academia and industry. However, recognition of Arabic handwritten and printed text presents difficu…
Arabic Handwritten Document OCR Solution with Binarization and Adaptive Scale Fusion Detection
Alhossien Waly, Bassant Tarek, Ali Feteha +3
The problem of converting images of text into plain text is a widely researched topic in both academia and industry. Arabic handwritten Text Recognation (AHTR) poses additional cha…