1 paper
Animesh Karnewar, Denis Korzhenkov, Amirhossein Habibian +1
Video Diffusion Transformers (DiTs) spend most of their compute inside the Self-Attention operation, whose cost grows quadratically, O(n2), with the number of latent t…