22 citations · 43 across the 3 of their papers we have counts for
3 papers
cs.CV2023★ 22 cited
VideoPoet: A Large Language Model for Zero-Shot Video Generation
Dan Kondratyuk, Lijun Yu, Xiuye Gu +28
We present VideoPoet, a language model capable of synthesizing high-quality video, with matching audio, from a large variety of conditioning signals. VideoPoet employs a decoder-on…
cs.CV2023
Diversity and Diffusion: Observations on Synthetic Image Distributions with Stable Diffusion
David Marwood, Shumeet Baluja, Yair Alon
Recent progress in text-to-image (TTI) systems, such as StableDiffusion, Imagen, and DALL-E 2, have made it possible to create realistic images with simple text prompts. It is temp…
cs.CV2020★ 21 cited
Wisdom of Committees: An Overlooked Approach To Faster and More Accurate Models
Xiaofang Wang, Dan Kondratyuk, Eric Christiansen +3
Committee-based models (ensembles or cascades) construct models by combining existing pre-trained ones. While ensembles and cascades are well-known techniques that were proposed be…