Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
CHASD: Language Increment-Calibrated Contrastive Decoding against Hallucination in LVLMs
Xiaoyi Huang, Kejia Zhang, Zhiming Luo
Large Vision-Language Models have shown strong multimodal reasoning capabilities, yet they remain susceptible to object hallucinations when language priors dominate insufficient or…
cs.CV2026
FastOCR: Dynamic Visual Fixation via KV Cache Pruning for Efficient Document Parsing
Zihan Tang, Leqi Shen, Hui Chen +7
Vision-Language Models (VLMs) have shown strong promise on Optical Character Recognition (OCR), yet the sheer number of visual tokens required to encode dense documents incurs proh…