1 paper
Abbas Mammadov, Jerry Y. Huang, Justin Lin +5
Reward fine-tuning aims to update a pre-trained flow-based generative model to improve the downstream reward of its generated samples. Existing methods typically formulate this pro…