Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
Taming Teacher Forcing for Masked Autoregressive Video Generation
Deyu Zhou, Quan Sun, Yuang Peng +8
We introduce MAGI, a hybrid video generation framework that combines masked modeling for intra-frame generation with causal modeling for next-frame generation. Our key innovation,…
cs.CV2023
Voila-A: Aligning Vision-Language Models with User's Gaze Attention
Kun Yan, Lei Ji, Zeyu Wang +3
In recent years, the integration of vision and language understanding has led to significant advancements in artificial intelligence, particularly through Vision-Language Models (V…