4 citations · 4 across the 2 of their papers we have counts for
2 papers
cs.CV2026
EditYourself: Audio-Driven Generation and Manipulation of Talking Head Videos with Diffusion Transformers
John Flynn, Wolfgang Paier, Dimitar Dinev +5
Current generative video models excel at producing novel content from text and image prompts, but leave a critical gap in editing existing pre-recorded videos, where minor alterati…
cs.CV2024★ 4 cited
StreamingT2V: Consistent, Dynamic, and Extendable Long Video Generation from Text
Roberto Henschel, Levon Khachatryan, Hayk Poghosyan +5
Text-to-video diffusion models enable the generation of high-quality videos that follow text instructions, making it easy to create diverse and individual content. However, existin…