2 papers
cs.CV2026
Compositional Video Generation via Inference-Time Guidance
Ariel Shaulov, Eitan Shaar, Amit Edenzon +2
Text-to-video diffusion models generate realistic videos, but often fail on prompts requiring fine-grained compositional understanding, such as relations between entities, attribut…
cs.CV2026
TokenTrim: Inference-Time Token Pruning for Autoregressive Long Video Generation
Ariel Shaulov, Eitan Shaar, Amit Edenzon +1
Auto-regressive video generation enables long video synthesis by iteratively conditioning each new batch of frames on previously generated content. However, recent work has shown t…