event-aware visual allocation 1gaussian mixture modeling 1keyframe selection 1long video understanding 1visual token budgeting 1
From the 1 of 3 linked papers with an AI index.
3 papers
cs.CV2026
Gaussian Mixture Modeling for Event-Aware Visual Allocation in Long Video Understanding
Yifan Lu, Ziqi Zhang, Chunfeng Yuan +3
The paper introduces GMM-EVA, a training-free framework that uses Gaussian Mixture Models to detect event-level structures in long videos and allocate visual tokens by selecting on…
cs.CV2026
IPAD-CLIP: Teaching CLIP to Detect Image Local Perceptual Artifacts
Juan Wang, Xinyu Sun, Ke Zhang +4
Current image quality assessment methods are heavily biased towards global distortions (e.g., noise, blur), neglecting local perceptual artifacts such as ghosting, lens flare, and…
cs.CV2025
Mitigating Hallucinations in Large Vision-Language Models by Self-Injecting Hallucinations
Yifan Lu, Ziqi Zhang, Chunfeng Yuan +5
Large Vision-Language Models (LVLMs) suffer from serious hallucination problems, where the model-generated responses are inconsistent with the visual inputs. Existing hallucination…