4 citations · 5 across the 2 of their papers we have counts for
2 papers
cs.LG2023★ 1 cited
Using Human Feedback to Fine-tune Diffusion Models without Any Reward Model
Kai Yang, Jian Tao, Jiafei Lyu +6
Using reinforcement learning with human feedback (RLHF) has shown significant promise in fine-tuning diffusion models. Previous methods start by training a reward model that aligns…
cs.AI2023★ 4 cited
Emergent collective intelligence from massive-agent cooperation and competition
Hanmo Chen, Stone Tao, Jiaxin Chen +6
Inspired by organisms evolving through cooperation and competition between different populations on Earth, we study the emergence of artificial collective intelligence through mass…