2 papers
cs.CV2025
World-To-Image: Grounding Text-to-Image Generation with Agent-Driven World Knowledge
Moo Hyun Son, Jintaek Oh, Sun Bin Mun +2
While text-to-image (T2I) models can synthesize high-quality images, their performance degrades significantly when prompted with novel or out-of-distribution (OOD) entities due to…
cs.CV2025
EO-VLM: VLM-Guided Energy Overload Attacks on Vision Models
Minjae Seo, Myoungsung You, Junhee Lee +4
Vision models are increasingly deployed in critical applications such as autonomous driving and CCTV monitoring, yet they remain susceptible to resource-consuming attacks. In this…