3 papers
cs.CV2024
Draw Like an Artist: Complex Scene Generation with Diffusion Model via Composition, Painting, and Retouching
Minghao Liu, Le Zhang, Yingjie Tian +3
Recent advances in text-to-image diffusion models have demonstrated impressive capabilities in image quality. However, complex scene generation remains relatively unexplored, and e…
cs.CV2024
Rethinking Video Segmentation with Masked Video Consistency: Did the Model Learn as Intended?
Chen Liang, Qiang Guo, Xiaochao Qu +2
Video segmentation aims at partitioning video sequences into meaningful segments based on objects or regions of interest within frames. Current video segmentation models are often…
cs.CV2024
TextMastero: Mastering High-Quality Scene Text Editing in Diverse Languages and Styles
Tong Wang, Xiaochao Qu, Ting Liu
Scene text editing aims to modify texts on images while maintaining the style of newly generated text similar to the original. Given an image, a target area, and target text, the t…