17 citations · 21 across the 3 of their papers we have counts for
3 papers
cs.CV2024★ 17 cited
Lumiere: A Space-Time Diffusion Model for Video Generation
Omer Bar-Tal, Hila Chefer, Omer Tov +14
We introduce Lumiere -- a text-to-video diffusion model designed for synthesizing videos that portray realistic, diverse and coherent motion -- a pivotal challenge in video synthes…
cs.CV2023★ 4 cited
Idempotent Generative Network
Assaf Shocher, Amil Dravid, Yossi Gandelsman +3
We propose a new approach for generative modeling based on training a neural network to be idempotent. An idempotent operator is one that can be applied sequentially without changi…
cs.CV2023
Teaching CLIP to Count to Ten
Roni Paiss, Ariel Ephrat, Omer Tov +4
Large vision-language models (VLMs), such as CLIP, learn rich joint image-text representations, facilitating advances in numerous downstream tasks, including zero-shot classificati…