◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Hao Wang

zhejiang university

22 papers hereh-index 18980 citations57 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author11
  • middle author8
  • last author2

Across the 21 of 22 papers where every author was matched, so the position is known.

fields
  • cs.LG12
  • cs.AI4
  • cs.CL3
  • cs.HC1
  • eess.SY1
  • stat.ML1
affiliations
  • zhejiang university
same name
  • Hao Wang — 23 papers, h 13
  • Hao Wang — 20 papers, h 7
  • Hao Wang — 19 papers, h 7
  • Hao Wang — 18 papers, h 5
  • Hao Wang — 17 papers, h 8
  • Hao Wang — 14 papers, h 3

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

works on
large language model evaluation 1preference learning 1query-only supervision 1rubric generation 1synthetic pairwise data 1

From the 1 of 22 linked papers with an AI index.

most citedOmniGAIA: Towards Native Omni-Modal AI Agents

1 citations · 1 across the 4 of their papers we have counts for

collaborators
Showing cs.CLShow all

3 papers · 1 filter

cs.CL2026

Rubrics on Trial: Evolving Rubrics from a Single Query via Synthetic Pairwise Evidence

Haocheng Yang, Licheng Pan, Xiaoxi Li +5

The paper proposes a query‑only method that automatically creates and validates fine‑grained rubrics for evaluating large language models by using synthetic rubric‑conditioned resp…

cs.CL2026

From Text to Talk: Audio-Language Model Needs Non-Autoregressive Joint Training

Tianqiao Liu, Xueyi Li, Hao Wang +4

Recent advances in large language models (LLMs) have attracted significant interest in extending their capabilities to multimodal scenarios, particularly for speech-to-speech conve…

cs.CL2026

ImplicitRM: Unbiased Reward Modeling from Implicit Preference Data for LLM alignment

Hao Wang, Haocheng Yang, Licheng Pan +7

Reward modeling represents a long-standing challenge in reinforcement learning from human feedback (RLHF) for aligning language models. Current reward modeling is heavily contingen…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.