1 paper · 1 filter
Zehang Wei, Jiaxin Dai, Jiamin Yan +1
While Multimodal Retrieval-Augmented Generation (M-RAG) enhances Large Vision-Language Models, it remains highly susceptible to cross-modal hallucinations, causal fabrications, and…