4 citations · 8 across the 11 of their papers we have counts for
1 paper · 1 filter
Kejia Zhang, Keda Tao, Zhiming Luo +3
Multimodal large language models (MLLMs) are prone to hallucinations, generating plausible but visually ungrounded outputs, partly because direct preference optimization (DPO) over…