1 paper
Myeongjun Oh, Gwangho Kim, Sungyoon Lee
Inference-time reward alignment steers pretrained diffusion and flow-based generative models to satisfy user-specified rewards without retraining. Recently, Sequential Monte Carlo…