1 citations · 1 across the 4 of their papers we have counts for
1 paper · 1 filter
Xuefeng Gao, Jiale Zha, Xun Yu Zhou
We propose a new reinforcement learning (RL) formulation for training continuous-time score-based diffusion models for generative AI to generate samples that maximize reward functi…