1 paper
Long Vu, Tan Ngo, Animesh Karnewar +5
Modern image-and-text-to-video diffusion models can synthesize highly realistic videos by iteratively denoising an initial Gaussian noise tensor conditioned on reference image and…