Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
When Does Visual Generation Help Visual Understanding in Unified Multimodal Models?
Yubo Zhu, Zhehan Kan, Jingyi Yang +6
Unified multimodal models (UMMs) can perform both understanding and generation, raising a central question: can visual generation improve understanding? Existing evaluations provid…
cs.CV2024
Detecting Dataset Abuse in Fine-Tuning Stable Diffusion Models for Text-to-Image Synthesis
Songrui Wang, Yubo Zhu, Wei Tong +1
Text-to-image synthesis has become highly popular for generating realistic and stylized images, often requiring fine-tuning generative models with domain-specific datasets for spec…