◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Yupei Yang

5 papers hereh-index 358 citations10 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author5

Across the 5 of 5 papers where every author was matched, so the position is known.

fields
  • cs.LG3
  • cs.CL1
  • cs.IR1

identity via Semantic Scholar / OpenAlex

activity
20242026
collaborators
Showing cs.LGShow all

3 papers · 1 filter

cs.LG2026

Factored Causal Representation Learning for Robust Reward Modeling in RLHF

Yupei Yang, Lin Yang, Wanxi Deng +5

A reliable reward model is essential for aligning large language models with human preferences through reinforcement learning from human feedback. However, standard reward models a…

cs.LG2024

Boosting Efficiency in Task-Agnostic Exploration through Causal Knowledge

Yupei Yang, Biwei Huang, Shikui Tu +1

The effectiveness of model training heavily relies on the quality of available training resources. However, budget constraints often impose limitations on data collection efforts.…

cs.LG2024

Towards Generalizable Reinforcement Learning via Causality-Guided Self-Adaptive Representations

Yupei Yang, Biwei Huang, Fan Feng +3

General intelligence requires quick adaption across tasks. While existing reinforcement learning (RL) methods have made progress in generalization, they typically assume only distr…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.