1 paper
Haopeng Li, Shitong Shao, Wenliang Zhong +4
Diffusion Transformers are fundamental for video and image generation, but their efficiency is bottlenecked by the quadratic complexity of attention. While block sparse attention a…