◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Sunli Chen

8 papers hereh-index 6187 citations8 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author5

Across the 5 of 8 papers where every author was matched, so the position is known.

fields
  • cs.CV4
  • cs.LG2
  • cs.AI1
  • cs.RO1

identity via Semantic Scholar / OpenAlex

activity
20232026
most citedHAZARD Challenge: Embodied Decision Making in Dynamically Changing Environments

2 citations · 2 across the 3 of their papers we have counts for

collaborators
Showing cs.LGShow all

2 papers · 1 filter

cs.LG2024

Improving Reinforcement Learning from Human Feedback with Efficient Reward Model Ensemble

Shun Zhang, Zhenfang Chen, Sunli Chen +3

Reinforcement Learning from Human Feedback (RLHF) is a widely adopted approach for aligning large language models with human values. However, RLHF relies on a reward model that is…

cs.LG2023

Iteratively Learn Diverse Strategies with State Distance Information

Wei Fu, Weihua Du, Jingwei Li +3

In complex reinforcement learning (RL) problems, policies with similar rewards may have substantially different behaviors. It remains a fundamental challenge to optimize rewards wh…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.