1 paper · 1 filter
Yifei Xia, Suhan Ling, Fangcheng Fu +4
Generating high-fidelity long videos with Diffusion Transformers (DiTs) is often hindered by significant latency, primarily due to the computational demands of attention mechanisms…