3 citations · 4 across the 4 of their papers we have counts for
4 papers
SmartControl: Enhancing ControlNet for Handling Rough Visual Conditions
Xiaoyu Liu, Yuxiang Wei, Ming Liu +4
Human visual imagination usually begins with analogies or rough sketches. For example, given an image with a girl playing guitar before a building, one may analogously imagine how…
VideoElevator: Elevating Video Generation Quality with Versatile Text-to-Image Diffusion Models
Yabo Zhang, Yuxiang Wei, Xianhui Lin +5
Text-to-image diffusion models (T2I) have demonstrated unprecedented capabilities in creating realistic and aesthetic images. On the contrary, text-to-video diffusion models (T2V)…
WordArt Designer API: User-Driven Artistic Typography Synthesis with Large Language Models on ModelScope
Jun-Yan He, Zhi-Qi Cheng, Chenyang Li +10
This paper introduces the WordArt Designer API, a novel framework for user-driven artistic typography synthesis utilizing Large Language Models (LLMs) on ModelScope. We address the…
VQ-Font: Few-Shot Font Generation with Structure-Aware Enhancement and Quantization
Mingshuai Yao, Yabo Zhang, Xianhui Lin +2
Few-shot font generation is challenging, as it needs to capture the fine-grained stroke styles from a limited set of reference glyphs, and then transfer to other characters, which…