Showing cs.CVShow all
3 papers · 1 filter
cs.CV2025
Rethinking Prompt Design for Inference-time Scaling in Text-to-Visual Generation
Subin Kim, Sangwoo Mo, Mamshad Nayeem Rizve +4
Achieving precise alignment between user intent and generated visuals remains a central challenge in text-to-visual generation, as a single attempt often fails to produce the desir…
cs.CV2025
FontAdapter: Instant Font Adaptation in Visual Text Generation
Myungkyu Koo, Subin Kim, Sangkyung Kwak +3
Text-to-image diffusion models have significantly improved the seamless integration of visual text into diverse image contexts. Recent approaches further improve control over font…
cs.CV2025
Tuning-Free Multi-Event Long Video Generation via Synchronized Coupled Sampling
Subin Kim, Seoung Wug Oh, Jui-Hsien Wang +2
While recent advancements in text-to-video diffusion models enable high-quality short video generation from a single prompt, generating real-world long videos in a single pass rema…