1 paper · 1 filter
Yaofu Liu, Wanli Lan, Jinxi Li +2
In DiT-based video generation models equipped with 3D Rotary Position Embeddings (3D RoPE), the attention mechanism remains a primary computational bottleneck due to its…