Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
Self-Aug: Query and Entropy Adaptive Decoding for Large Vision-Language Models
Eun Woo Im, Muhammad Kashif Ali, Vivek Gupta
Large Vision-Language Models (LVLMs) have demonstrated remarkable multimodal capabilities, but they inherit the tendency to hallucinate from their underlying language models. While…
cs.CV2025
VidHalluc: Evaluating Temporal Hallucinations in Multimodal Large Language Models for Video Understanding
Chaoyu Li, Eun Woo Im, Pooyan Fazli
Multimodal large language models (MLLMs) have recently shown significant advancements in video understanding, excelling in content reasoning and instruction-following tasks. Howeve…