1 paper
Ruiqi Lai, Dakai An, Wei Gao +6
Reinforcement learning (RL) post-training of Diffusion Transformers (DiTs) is prohibitively expensive, requiring thousands of high-end GPUs. Existing works explore two directions t…