◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

O. Mohammed

4 papers hereh-index 62k citations10 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author3

Across the 3 of 4 papers where every author was matched, so the position is known.

fields
  • cs.CV3
  • cs.CL1

identity via Semantic Scholar / OpenAlex

activity
20212023
most citedVLMo: Unified Vision-Language Pre-Training with Mixture-of-Modality-Experts

288 citations · 605 across the 4 of their papers we have counts for

collaborators

4 papers

cs.CV2023★ 2 cited

ArK: Augmented Reality with Knowledge Interactive Emergent Ability

Qiuyuan Huang, Jae Sung Park, Abhinav Gupta +8

Despite the growing adoption of mixed reality and interactive AI agents, it remains challenging for these systems to generate high quality 2D/3D scenes in unseen environments. The…

cs.CL2023★ 164 cited

Language Is Not All You Need: Aligning Perception with Language Models

Shaohan Huang, Li Dong, Wenhui Wang +15

A big convergence of language, multimodal perception, action, and world modeling is a key step toward artificial general intelligence. In this work, we introduce Kosmos-1, a Multim…

cs.CV2022★ 151 cited

Image as a Foreign Language: BEiT Pretraining for All Vision and Vision-Language Tasks

Wenhui Wang, Hangbo Bao, Li Dong +8

A big convergence of language, vision, and multimodal pretraining is emerging. In this work, we introduce a general-purpose multimodal foundation model BEiT-3, which achieves state…

cs.CV2021★ 288 cited

VLMo: Unified Vision-Language Pre-Training with Mixture-of-Modality-Experts

Hangbo Bao, Wenhui Wang, Li Dong +5

We present a unified Vision-Language pretrained Model (VLMo) that jointly learns a dual encoder and a fusion encoder with a modular Transformer network. Specifically, we introduce…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.