2 papers
cs.CV2025
EVLM: Self-Reflective Multimodal Reasoning for Cross-Dimensional Visual Editing
Umar Khalid, Kashif Munir, Hasan Iqbal +6
Editing complex visual content from ambiguous or partially specified instructions remains a core challenge in vision-language modeling. Existing models can contextualize content bu…
cs.LG2025
On Transfer-based Universal Attacks in Pure Black-box Setting
Mohammad A. A. K. Jalwana, Naveed Akhtar, Ajmal Mian +2
Despite their impressive performance, deep visual models are susceptible to transferable black-box adversarial attacks. Principally, these attacks craft perturbations in a target m…