Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
Manipulating Multimodal Agents via Cross-Modal Prompt Injection
Le Wang, Zonghao Ying, Tianyuan Zhang +5
The emergence of multimodal large language models has redefined the agent paradigm by integrating language and vision modalities with external data sources, enabling agents to bett…
cs.CV2025
CogMorph: Cognitive Morphing Attacks for Text-to-Image Models
Zonglei Jing, Zonghao Ying, Le Wang +4
The development of text-to-image (T2I) generative models, that enable the creation of high-quality synthetic images from textual prompts, has opened new frontiers in creative desig…