Showing cs.CVShow all
2 papers · 1 filter
cs.CV2024
Video Diffusion Models: A Survey
Andrew Melnik, Michal Ljubljanac, Cong Lu +3
Diffusion generative models have recently become a powerful technique for creating and modifying high-quality, coherent video content. This survey provides a comprehensive overview…
cs.CV2024
Pre-trained Text-to-Image Diffusion Models Are Versatile Representation Learners for Control
Gunshi Gupta, Karmesh Yadav, Yarin Gal +4
Embodied AI agents require a fine-grained understanding of the physical world mediated through visual and language inputs. Such capabilities are difficult to learn solely from task…