2 papers
cs.CV2026
From Clouds to Hallucinations: Atmospheric Retrieval Hijacking in Remote Sensing Vision-Language RAG
Jiaju Han, Chao Li, Chengyin Hu +8
Multimodal RAG systems increasingly rely on vision-language retrievers to ground visual queries in external textual evidence. Existing adversarial studies on RAG mainly manipulate…
cs.CV2026
Towards multi-modal forgery representation learning for AI-generated video detection and localization
Dat Le, Khoa Nguyen, Xin Wang +1
Recent advances in generative AI have democratized video creation at scale. AI-generated videos, including partially manipulated clips across visual and audio channels, pose escala…