1 paper
Luke Budny, Yuhong Guo, Kevin Cheung
Modern text-to-image diffusion models, such as diffusion transformers (DiT), rely on timestep or prompt embeddings to modulate the strength of the denoising process in each timeste…