◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Yong Cheng

5 papers hereh-index 7286 citations8 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author2
  • middle author3

Across the 5 of 5 papers where every author was matched, so the position is known.

fields
  • cs.CL4
  • cs.CV1
same name
  • Yong Cheng — 9 papers, h 7
  • Yong Cheng — 8 papers, h 16
  • Yong Cheng — 8 papers, h 51
  • Yong Cheng — 8 papers, h 3
  • Yong Cheng — 5 papers, h 3
  • Yong Cheng — 5 papers, h 4

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

most citedmSLAM: Massively multilingual joint pre-training for speech and text

59 citations · 73 across the 5 of their papers we have counts for

collaborators
Showing cs.CLShow all

4 papers · 1 filter

cs.CL2022★ 5 cited

Mu2SLAM: Multitask, Multilingual Speech and Language Models

Yong Cheng, Yu Zhang, Melvin Johnson +2

We present Mu2SLAM, a multilingual sequence-to-sequence model pre-trained jointly on unlabeled speech, unlabeled text and supervised data spanning Automatic Speech Recognition…

cs.CL2022

Multilingual Mix: Example Interpolation Improves Multilingual Neural Machine Translation

Yong Cheng, Ankur Bapna, Orhan Firat +3

Multilingual neural machine translation models are trained to maximize the likelihood of a mix of examples drawn from multiple language pairs. The dominant inductive bias applied t…

cs.CL2022★ 2 cited

Examining Scaling and Transfer of Language Model Architectures for Machine Translation

Biao Zhang, Behrooz Ghorbani, Ankur Bapna +4

Natural language understanding and generation models follow one of the two dominant architectural paradigms: language models (LMs) that process concatenated sequences in a single s…

cs.CL2022★ 59 cited

mSLAM: Massively multilingual joint pre-training for speech and text

Ankur Bapna, Colin Cherry, Yu Zhang +6

We present mSLAM, a multilingual Speech and LAnguage Model that learns cross-lingual cross-modal representations of speech and text by pre-training jointly on large amounts of unla…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.