◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Yuxin Wu

3 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author3

Across the 3 of 3 papers where every author was matched, so the position is known.

fields
  • cs.LG3
same name
  • Yuxin Wu — 6 papers, h 26
  • Yuxin Wu — 3 papers
  • Yuxin Wu — 3 papers
  • Yuxin Wu — 2 papers
  • Yuxin Wu — 2 papers
  • Yuxin Wu — 1 paper

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

most citedMuon is Scalable for LLM Training

1 citations · 1 across the 1 of their papers we have counts for

collaborators

3 papers

cs.LG2025★ 1 cited

Muon is Scalable for LLM Training

Jingyuan Liu, Jianlin Su, Xingcheng Yao +25

Recently, the Muon optimizer based on matrix orthogonalization has demonstrated strong results in training small-scale language models, but the scalability to larger models has not…

cs.LG2025

MoBA: Mixture of Block Attention for Long-Context LLMs

Enzhe Lu, Zhejun Jiang, Jingyuan Liu +22

Scaling the effective context length is essential for advancing large language models (LLMs) toward artificial general intelligence (AGI). However, the quadratic increase in comput…

cs.LG2018

Natural Environment Benchmarks for Reinforcement Learning

Amy Zhang, Yuxin Wu, Joelle Pineau

While current benchmark reinforcement learning (RL) tasks have been useful to drive progress in the field, they are in many ways poor substitutes for learning with real-world data.…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.