◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Wayne Xin Zhao

6 papers hereh-index 345 citations7 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author6

Across the 6 of 6 papers where every author was matched, so the position is known.

fields
  • cs.CL4
  • cs.LG2
same name
  • Wayne Xin Zhao — 108 papers, h 72
  • Wayne Xin Zhao — 17 papers, h 8
  • Wayne Xin Zhao — 11 papers, h 6
  • Wayne Xin Zhao — 9 papers, h 4
  • Wayne Xin Zhao — 8 papers, h 4
  • Wayne Xin Zhao — 4 papers, h 3

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

most citedRethinking Sample Polarity in Reinforcement Learning with Verifiable Rewards

1 citations · 1 across the 6 of their papers we have counts for

collaborators
Showing cs.LGShow all

2 papers · 1 filter

cs.LG2026

MergeMix: Optimizing Mid-Training Data Mixtures via Learnable Model Merging

Jiapeng Wang, Changxin Tian, Kunlong Chen +5

Optimizing data mixtures is essential for unlocking the full potential of large language models (LLMs), yet identifying the optimal composition remains computationally prohibitive…

cs.LG2025

Towards High Data Efficiency in Reinforcement Learning with Verifiable Reward

Xinyu Tang, Zhenduo Zhang, Yurou Liu +4

Recent advances in large reasoning models have leveraged reinforcement learning with verifiable rewards (RLVR) to improve reasoning capabilities. However, scaling these methods typ…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.