2 citations · 2 across the 3 of their papers we have counts for
3 papers
AutoArabic: A Three-Stage Framework for Localizing Video-Text Retrieval Benchmarks
Mohamed Eltahir, Osamah Sarraj, Abdulrahman Alfrihidi +4
Video-to-text and text-to-video retrieval are dominated by English benchmarks (e.g. DiDeMo, MSR-VTT) and recent multilingual corpora (e.g. RUDDER), yet Arabic remains underserved,…
Multimodal Lengthy Videos Retrieval Framework and Evaluation Metric
Mohamed Eltahir, Osamah Sarraj, Mohammed Bremoo +5
Precise video retrieval requires multi-modal correlations to handle unseen vocabulary and scenes, becoming more complex for lengthy videos where models must perform effectively wit…
Local and Global Contextual Features Fusion for Pedestrian Intention Prediction
Mohsen Azarmi, Mahdi Rezaei, Tanveer Hussain +1
Autonomous vehicles (AVs) are becoming an indispensable part of future transportation. However, safety challenges and lack of reliability limit their real-world deployment. Towards…