1 paper
Yao Teng, Minxuan Lin, Xian Liu +3
Recent video generation models largely rely on video autoencoders that compress pixel-space videos into latent representations. However, existing video autoencoders suffer from thr…