3 papers
cs.CV2026
Adaptive 1D Video Diffusion Autoencoder
Yao Teng, Minxuan Lin, Xian Liu +3
Recent video generation models largely rely on video autoencoders that compress pixel-space videos into latent representations. However, existing video autoencoders suffer from thr…
cs.CV2026
FSVideo: Fast Speed Video Diffusion Model in a Highly-Compressed Latent Space
FSVideo Team, Qingyu Chen, Zhiyuan Fang +17
We introduce FSVideo, a fast speed transformer-based image-to-video (I2V) diffusion framework. We build our framework on the following key components: 1.) a new video autoencoder w…
cs.CV2025
IP-Prompter: Training-Free Theme-Specific Image Generation via Dynamic Visual Prompting
Yuxin Zhang, Minyan Luo, Weiming Dong +6
The stories and characters that captivate us as we grow up shape unique fantasy worlds, with images serving as the primary medium for visually experiencing these realms. Personaliz…