4 citations · 4 across the 2 of their papers we have counts for
2 papers
cs.MM2024
Investigating Conceptual Blending of a Diffusion Model for Improving Nonword-to-Image Generation
Chihaya Matsuhira, Marc A. Kastner, Takahiro Komamizu +2
Text-to-image diffusion models sometimes depict blended concepts in the generated images. One promising use case of this effect would be the nonword-to-image generation task which…
cs.MM2023★ 4 cited
IPA-CLIP: Integrating Phonetic Priors into Vision and Language Pretraining
Chihaya Matsuhira, Marc A. Kastner, Takahiro Komamizu +4
Recently, large-scale Vision and Language (V\&L) pretraining has become the standard backbone of many multimedia systems. While it has shown remarkable performance even in unseen s…