2 citations · 3 across the 2 of their papers we have counts for
2 papers
cs.CV2024★ 1 cited
Animated Stickers: Bringing Stickers to Life with Video Diffusion
David Yan, Winnie Zhang, Luxin Zhang +15
We introduce animated stickers, a video diffusion model which generates an animation conditioned on a text prompt and static sticker image. Our model is built on top of the state-o…
cs.CV2023★ 2 cited
DISGO: Automatic End-to-End Evaluation for Scene Text OCR
Mei-Yuh Hwang, Yangyang Shi, Ankit Ramchandani +6
This paper discusses the challenges of optical character recognition (OCR) on natural scenes, which is harder than OCR on documents due to the wild content and various image backgr…