2 papers
cs.LG2025
Results of the NeurIPS 2023 Neural MMO Competition on Multi-task Reinforcement Learning
Joseph Suárez, Kyoung Whan Choe, David Bloomin +22
We present the results of the NeurIPS 2023 Neural MMO Competition, which attracted over 200 participants and submissions. Participants trained goal-conditional policies that genera…
cs.LG2024
Using Human Feedback to Fine-tune Diffusion Models without Any Reward Model
Kai Yang, Jian Tao, Jiafei Lyu +6
Using reinforcement learning with human feedback (RLHF) has shown significant promise in fine-tuning diffusion models. Previous methods start by training a reward model that aligns…