7 papers
Multimodal Markup Document Models for Graphic Design Completion
Kotaro Kikuchi, Ukyo Honda, Naoto Inoue +3
We introduce MarkupDM, a multimodal markup document model that represents graphic design as an interleaved multimodal document consisting of both markup language and images. Unlike…
Automatic Text Box Placement for Supporting Typographic Design
Jun Muraoka, Daichi Haraguchi, Naoto Inoue +3
In layout design for advertisements and web pages, balancing visual appeal and communication efficiency is crucial. This study examines automated text box placement in incomplete l…
OTR: Synthesizing Overlay Text Dataset for Text Removal
Jan Zdenek, Wataru Shimoda, Kota Yamaguchi
Text removal is a crucial task in computer vision with applications such as privacy preservation, image editing, and media reuse. While existing research has primarily focused on s…
OnomatoGen: Onomatopoeia Generation with the Alpha-Channel in Manga
Takara Taniguchi, Wataru Shimoda, Kota Yamaguchi +1
Onomatopoeia is an important element for textual messaging in manga. Unlike character dialogue in manga, onomatopoeic expressions are visually stylized, with variations in shape, s…
Total Disentanglement of Font Images into Style and Character Class Features
Daichi Haraguchi, Wataru Shimoda, Kota Yamaguchi +1
In this paper, we demonstrate a total disentanglement of font images. Total disentanglement is a neural network-based method for decomposing each font image nonlinearly and complet…
Type-R: Automatically Retouching Typos for Text-to-Image Generation
Wataru Shimoda, Naoto Inoue, Daichi Haraguchi +3
While recent text-to-image models can generate photorealistic images from text prompts that reflect detailed instructions, they still face significant challenges in accurately rend…