Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
High-Fidelity Text-to-Image Generation from Pre-Trained Vision-Language Models via Distribution-Conditioned Diffusion Decoding
Ji Woo Hong, Hee Suk Yoon, Gwanhyeong Koo +5
Recent large-scale vision-language models (VLMs) have shown remarkable text-to-image generation capabilities, yet their visual fidelity remains constrained by the discrete image to…
cs.CV2025
E-MD3C: Taming Masked Diffusion Transformers for Efficient Zero-Shot Object Customization
Trung X. Pham, Zhang Kang, Ji Woo Hong +2
We propose E-MD3C (fficient asked iffusion Transformer with Disentangled onditions and ompact $\underline…