1 paper · 1 filter
Linfeng Fan, Yuan Tian, Ziwei Li +1
Video Large Language Models (Video-LLMs) remain prone to spatiotemporal hallucinations, often generating visually unsupported details or incorrect temporal relations. Existing miti…