◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Yang Gan

7 papers hereh-index 346 citations8 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author7

Across the 7 of 7 papers where every author was matched, so the position is known.

fields
  • cs.CL4
  • cs.CV2
  • cs.DC1

identity via Semantic Scholar / OpenAlex

activity
20242026
collaborators
Showing cs.CLShow all

4 papers · 1 filter

cs.CL2026

GRAM: A Generative Foundation Reward Model for Reward Generalization

Chenglong Wang, Yang Gan, Yifu Huo +8

In aligning large language models (LLMs), reward models have played an important role, but are standardly trained as discriminative models and rely only on labeled human preference…

cs.CL2025

Probing Preference Representations: A Multi-Dimensional Evaluation and Analysis Method for Reward Models

Chenglong Wang, Yifu Huo, Yang Gan +10

Previous methods evaluate reward models by testing them on a fixed pairwise ranking test set, but they typically do not provide performance information on each preference dimension…

cs.CL2025

MRO: Enhancing Reasoning in Diffusion Language Models via Multi-Reward Optimization

Chenglong Wang, Yang Gan, Hang Zhou +10

Recent advances in diffusion language models (DLMs) have presented a promising alternative to traditional autoregressive large language models (LLMs). However, DLMs still lag behin…

cs.CL2024

LRHP: Learning Representations for Human Preferences via Preference Pairs

Chenglong Wang, Yang Gan, Yifu Huo +7

To improve human-preference alignment training, current research has developed numerous preference datasets consisting of preference pairs labeled as "preferred" or "dispreferred".…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.