◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Hua Yang

9 papers hereh-index 4141 citations11 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author1
  • middle author5
  • last author2

Across the 8 of 9 papers where every author was matched, so the position is known.

fields
  • cs.CV6
  • cs.RO2
  • cs.CL1
same name
  • Hua Yang — 4 papers, h 2
  • Hua Yang — 3 papers, h 2
  • Hua Yang — 2 papers, h 1
  • Hua Yang — 1 paper, h 2
  • Hua Yang — 1 paper, h 2
  • Hua Yang — 1 paper, h 5

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

works on
language models 1open-ended generation 1policy optimization 1reward modeling 1self-evolving rubrics 1test-time reinforcement learning 1

From the 1 of 9 linked papers with an AI index.

collaborators
Showing cs.CLShow all

2 papers · 1 filter

cs.CL2026

SERPO: Self-Evolving Rubric Policy Optimization for Open-Ended Test-Time Reinforcement Learning

Jianze Wang, Kunwang Zheng, Ying Liu +5

The paper introduces SERPO, a test-time reinforcement learning approach that lets language models self‑improve during inference by jointly evolving response evidence, query‑specifi…

cs.CL2024

MCSD: An Efficient Language Model with Diverse Fusion

Hua Yang, Duohai Li, Shiman Li

Transformers excel in Natural Language Processing (NLP) due to their prowess in capturing long-term dependencies but suffer from exponential resource consumption with increasing se…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.