Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
Mitigating Compounding Error via Video Representation Regularization
Taiye Chen, Qi Zhang, Yisen Wang
Video diffusion-based world models enable long autoregressive video generation for robotics, autonomous driving and simulation tasks, yet sliding-window autoregressive inference su…
cs.CV2025
Recurrent Autoregressive Diffusion: Global Memory Meets Local Attention
Taiye Chen, Zihan Ding, Anjian Li +4
Recent advancements in video generation has shifted from bidirectional models for short videos to autoregressive ones for ultra long video generation. Previous models, which usuall…
cs.CV2025
VRAG: Learning World Models for Interactive Video Generation
Taiye Chen, Xun Hu, Zihan Ding +1
Foundational world models must be both interactive and preserve spatiotemporal coherence for effective future planning with action choices. However, present models for long video g…