4 papers
Adaptive Hybrid Caching for Efficient Text-to-Video Diffusion Model Acceleration
Yuanxin Wei, Lansong Diao, Bujiao Chen +6
Efficient video generation models are increasingly vital for multimedia synthetic content generation. Leveraging the Transformer architecture and the diffusion process, video DiT m…
SRDiffusion: Accelerate Video Diffusion Inference via Sketching-Rendering Cooperation
Shenggan Cheng, Yuanxin Wei, Lansong Diao +8
Leveraging the diffusion transformer (DiT) architecture, models like Sora, CogVideoX and Wan have achieved remarkable progress in text-to-video, image-to-video, and video editing t…
Wan: Open and Advanced Large-Scale Video Generative Models
Team Wan, Ang Wang, Baole Ai +58
This report presents Wan, a comprehensive and open suite of video foundation models designed to push the boundaries of video generation. Built upon the mainstream diffusion transfo…
Static Batching of Irregular Workloads on GPUs: Framework and Application to Efficient MoE Model Inference
Yinghan Li, Yifei Li, Jiejing Zhang +13
It has long been a problem to arrange and execute irregular workloads on massively parallel devices. We propose a general framework for statically batching irregular workloads into…