1 paper
Minh Vu, Geigh Zollicoffer, Huy Mai +3
Multimodal Machine Learning systems, particularly those aligning text and image data like CLIP/BLIP models, have become increasingly prevalent, yet remain susceptible to adversaria…