1 paper · 1 filter
Jianfeng Cai, Wengang Zhou, Zongmeng Zhang +3
Multimodal large language models (MLLMs) have achieved remarkable progress in video understanding.However, hallucination, where the model generates plausible yet incorrect outputs,…