1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.CY2025
OpenReview Should be Protected and Leveraged as a Community Asset for Research in the Era of Large Language Models
Hao Sun, Yunyi Shen, Mihaela van der Schaar
In the era of large language models (LLMs), high-quality, domain-rich, and continuously evolving datasets capturing expert-level knowledge, core human values, and reasoning are inc…
cs.CL2025★ 1 cited
Reusing Embeddings: Reproducible Reward Model Research in Large Language Model Alignment without GPUs
Hao Sun, Yunyi Shen, Jean-Francois Ton +1
Large Language Models (LLMs) have made substantial strides in structured tasks through Reinforcement Learning (RL), demonstrating proficiency in mathematical reasoning and code gen…
cs.CL2025
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment
Yunyi Shen, Hao Sun, Jean-François Ton
Building neural reward models from human preferences is a pivotal component in reinforcement learning from human feedback (RLHF) and large language model alignment research. Given…