1 paper · 1 filter
Eun Woo Im, Muhammad Kashif Ali, Vivek Gupta
Large Vision-Language Models (LVLMs) have demonstrated remarkable multimodal capabilities, but they inherit the tendency to hallucinate from their underlying language models. While…