1 paper · 1 filter
Hyeonjeong Ha, Qiusi Zhan, Jeonghwan Kim +6
Retrieval-augmented generation (RAG) has become a common practice in multimodal large language models (MLLM) to enhance factual grounding and reduce hallucination. Yet, its relianc…