3 papers
cs.CV2026
Wiener Representation Filtering for VLM Hallucination Suppression
Ameen Ali, Tamim Zoabi, Lidor Brami +1
Vision-language models (VLMs) excel at open-ended captioning and visual QA but often describe objects, attributes, or relations absent from the image, a phenomenon known as object…
cs.LG2026
Mean-Field Parallel Decoding for Discrete Diffusion Language Models
Tamim Zoabi, Ameen Ali, Liran Ringel +1
Discrete diffusion language models enable parallel token generation, offering a pathway to low-latency decoding. However, selecting tokens independently by marginal confidence limi…
cs.CV2025
Suppressing VLM Hallucinations with Spectral Representation Filtering
Ameen Ali, Tamim Zoabi, Lior Wolf
Vision-language models (VLMs) frequently produce hallucinations in the form of descriptions of objects, attributes, or relations that do not exist in the image due to over-reliance…