1 paper · 1 filter
Chengzhi Yu, Yifan Xu, Yifan Chen +1
Recently, large vision-language models (LVLMs) have risen to be a promising approach for multimodal tasks. However, principled hallucination mitigation remains a critical challenge…