◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Abdelrahman Mohamed

4 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author1
  • last author1

Across the 2 of 4 papers where every author was matched, so the position is known.

fields
  • cs.CL2
  • cs.CV1
  • eess.AS1
same name
  • Abdelrahman Mohamed — 2 papers
  • Abdelrahman Mohamed — 1 paper

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20222025
most citedLearning Audio-Visual Speech Representation by Masked Multimodal Cluster Prediction

113 citations · 113 across the 4 of their papers we have counts for

collaborators

4 papers

cs.CL2025

JEEM: Vision-Language Understanding in Four Arabic Dialects

Karima Kadaoui, Hanin Atwany, Hamdan Al-Ali +7

We introduce JEEM, a benchmark designed to evaluate Vision-Language Models (VLMs) on visual understanding across four Arabic-speaking countries: Jordan, The Emirates, Egypt, and Mo…

cs.CV2023

Violet: A Vision-Language Model for Arabic Image Captioning with Gemini Decoder

Abdelrahman Mohamed, Fakhraddin Alwajih, El Moatez Billah Nagoudi +2

Although image captioning has a vast array of applications, it has not reached its full potential in languages other than English. Arabic, for instance, although the native languag…

cs.CL2022

STOP: A dataset for Spoken Task Oriented Semantic Parsing

Paden Tomasello, Akshat Shrivastava, Daniel Lazar +12

End-to-end spoken language understanding (SLU) predicts intent directly from audio using a single model. It promises to improve the performance of assistant systems by leveraging a…

eess.AS2022★ 113 cited

Learning Audio-Visual Speech Representation by Masked Multimodal Cluster Prediction

Bowen Shi, Wei-Ning Hsu, Kushal Lakhotia +1

Video recordings of speech contain correlated audio and visual information, providing a strong signal for speech representation learning from the speaker's lip movements and the pr…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.