Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
EVL-MCoT: Enhanced Vision-Language Multi-CoT for Harmful Meme Detection
Hao Yang, Jin Wang, Xuejie Zhang
MEMEs are widely used on the internet and often carry strong elements of sarcasm or irony. Understanding their hidden meanings typically requires a joint interpretation of text and…
cs.CV2025
Vision-aware Multimodal Prompt Tuning for Uploadable Multi-source Few-shot Domain Adaptation
Kuanghong Liu, Jin Wang, Kangjian He +2
Conventional multi-source domain few-shot adaptation (MFDA) faces the challenge of further reducing the load on edge-side devices in low-resource scenarios. Considering the native…