1 citations · 1 across the 6 of their papers we have counts for
1 paper · 1 filter
Zhe Gao, Shiyu Shen, Taifeng Chai +7
Existing Multimodal Large Language Models (MLLMs) often suffer from hallucinations in long video understanding (LVU), primarily due to the imbalance between textual and visual toke…