◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Han Wang

3 papers hereh-index 3137 citations10 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author2
  • middle author1

Across the 3 of 3 papers where every author was matched, so the position is known.

fields
  • cs.LG2
  • cs.CL1
same name
  • Han Wang — 22 papers
  • Han Wang — 18 papers
  • Han Wang — 9 papers
  • Han Wang — 7 papers
  • Han Wang — 7 papers
  • Han Wang — 6 papers

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

collaborators

3 papers

cs.CL2025

Text2Grad: Reinforcement Learning from Natural Language Feedback

Hanyang Wang, Lu Wang, Chaoyun Zhang +5

Traditional RLHF optimizes language models with coarse, scalar rewards that mask the fine-grained reasons behind success or failure, leading to slow and opaque learning. Recent wor…

cs.LG2025

Fine-Tuning without Performance Degradation

Han Wang, Adam White, Martha White

Fine-tuning policies learned offline remains a major challenge in application domains. Monotonic performance improvement during \emph{fine-tuning} is often challenging, as agents t…

cs.LG2025

Fat-to-Thin Policy Optimization: Offline RL with Sparse Policies

Lingwei Zhu, Han Wang, Yukie Nagai

Sparse continuous policies are distributions that can choose some actions at random yet keep strictly zero probability for the other actions, which are radically different from the…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.