1 paper
Tuna Tuncer, Felix Becker, Thomas Pfeil
Chunk-wise autoregressive video diffusion models rely on a KV cache of previously generated chunks to avoid redundant computation, but this cache quickly becomes a memory bottlenec…