3 papers
cs.CV2024
Type-R: Automatically Retouching Typos for Text-to-Image Generation
Wataru Shimoda, Naoto Inoue, Daichi Haraguchi +3
While recent text-to-image models can generate photorealistic images from text prompts that reflect detailed instructions, they still face significant challenges in accurately rend…
cs.CV2024
Can GPTs Evaluate Graphic Design Based on Design Principles?
Daichi Haraguchi, Naoto Inoue, Wataru Shimoda +3
Recent advancements in foundation models show promising capability in graphic design generation. Several studies have started employing Large Multimodal Models (LMMs) to evaluate g…
cs.CV2023
Selective Scene Text Removal
Hayato Mitani, Akisato Kimura, Seiichi Uchida
Scene text removal (STR) is the image transformation task to remove text regions in scene images. The conventional STR methods remove all scene text. This means that the existing m…