1 paper
Archer Moore, Mingming Gong, Liam Hodgkinson
Reinforcement learning from human feedback (RLHF) for 3D generation is now established across a number of works, but most existing pipelines optimise explicit surface representatio…