1 paper
Shanghao Liu, Xiaoyun Yu, Wanting Li +1
Attention computation makes inference expensive in video diffusion transformers (vDiTs), which generate videos through iterative denoising. Block-sparse attention (BSA) reduces thi…