6 citations · 6 across the 10 of their papers we have counts for
1 paper · 2 filters
Feiran Zhang, Yixin Wu, Zhenghua Wang +4
Vision-Language Models (VLMs) have demonstrated remarkable progress in multimodal tasks, but remain susceptible to hallucinations, where generated text deviates from the underlying…