activity
20242026
collaborators

18 papers

cs.CV2026

MaineCoon: Pursuing A Real-Time Audio-Visual Social World Model

Lichen Bai, Tianhao Zhang, Shitong Shao +14

As an increasing majority of global video content is consumed on social platforms for interactive social purposes, video generation models built for social worlds are important but…

cs.CV2026

Optimizing Few-Step Generation with Adaptive Matching Distillation

Lichen Bai, Zikai Zhou, Shitong Shao +5

Distribution Matching Distillation (DMD) is a powerful acceleration paradigm, yet its stability is often compromised in Forbidden Zone, regions where the real teacher provides unre…

cs.CV2026

LIVEditor-14B: Lightning Unified Video Editing via In-Context Sparse Attention

Shitong Shao, Zikai Zhou, Haopeng Li +4

Video editing has evolved toward In-Context Learning (ICL) paradigms, yet the resulting quadratic attention costs create a critical computational bottleneck. In this work, we propo…

cs.CV2026

Exploring Data-Free LoRA Transferability for Video Diffusion Models

Yuchen Wang, Wenliang Zhong, Lichen Bai +6

Video diffusion models leveraging step distillation or causal distillation have achieved remarkable performance. However, adapting existing LoRAs to these variants remains a critic…

cs.CV2026

Reflective Flow Sampling Enhancement

Zikai Zhou, Muyao Wang, Shitong Shao +4

The growing demand for text-to-image generation has led to rapid advances in generative modeling. Recently, text-to-image diffusion models trained with flow matching algorithms, su…

cs.CV2026

Efficient Video Diffusion Models: Advancements and Challenges

Shitong Shao, Lichen Bai, Pengfei Wan +2

Video diffusion models have rapidly become the dominant paradigm for high-fidelity generative video synthesis, but their practical deployment remains constrained by severe inferenc…