◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Yu Cheng

19 papers hereh-index 5122 citations20 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author9
  • last author9

Across the 18 of 19 papers where every author was matched, so the position is known.

fields
  • cs.AI8
  • cs.CL5
  • cs.LG4
  • cs.CV2
same name
  • Yu Cheng — 15 papers, h 6
  • Yu Cheng — 14 papers, h 10
  • Yu Cheng — 14 papers, h 6
  • Yu Cheng — 14 papers, h 4
  • Yu Cheng — 13 papers, h 10
  • Yu Cheng — 10 papers, h 6

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

collaborators
Showing cs.LGShow all

4 papers · 1 filter

cs.LG2026

How Does Reasoning Flow? Tracing Attention-Induced Information Flow for Targeted RL in LLMs

Zhichen Dong, Yang Li, Yuhan Sun +9

Token-level credit assignment remains a key obstacle for reinforcement learning (RL) in large language models (LLMs), where RL recipes typically treat all tokens equally, failing t…

cs.LG2026

The Cancellation Hypothesis in Critic-Free RL: From Outcome Rewards to Token Credits

Tianhao Cheng, Zeyu Huang, Zihan Qiu +5

A commonly accepted explanation of critic-free RL for LLMs, based on sequence-level rewards, is that it reinforces successful rollouts with a positive advantage while penalizing fa…

cs.LG2026

ExGRPO: Learning to Reason from Experience

Runzhe Zhan, Yafu Li, Zhi Wang +5

Reinforcement learning from verifiable rewards (RLVR) is an emerging paradigm for improving the reasoning ability of large language models. However, standard on-policy training dis…

cs.LG2025★ 1 cited

Rethinking Entropy Regularization in Large Reasoning Models

Yuxian Jiang, Yafu Li, Guanxu Chen +3

Reinforcement learning with verifiable rewards (RLVR) has shown great promise in enhancing the reasoning abilities of large reasoning models (LRMs). However, it suffers from a crit…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.