3 citations · 3 across the 2 of their papers we have counts for
2 papers
cs.HC2025★ 3 cited
ThematicPlane: Bridging Tacit User Intent and Latent Spaces for Image Generation
Daniel Lee, Nikhil Sharma, Donghoon Shin +4
Generative AI has made image creation more accessible, yet aligning outputs with nuanced creative intent remains challenging, particularly for non-experts. Existing tools often req…
cs.AI2024
ARMADA: Attribute-Based Multimodal Data Augmentation
Xiaomeng Jin, Jeonghwan Kim, Yu Zhou +4
In Multimodal Language Models (MLMs), the cost of manually annotating high-quality image-text pair data for fine-tuning and alignment is extremely high. While existing multimodal d…