1 citations · 1 across the 3 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
PAS : Prelim Attention Score for Detecting Object Hallucinations in Large Vision--Language Models
Nhat Hoang-Xuan, Minh Vu, My T. Thai +1
Large vision-language models (LVLMs) are powerful, yet they remain unreliable due to object hallucinations. In this work, we show that in many hallucinatory predictions the LVLM ef…
cs.CV2024
Patchfinder: Leveraging Visual Language Models for Accurate Information Retrieval using Model Uncertainty
Roman Colman, Minh Vu, Manish Bhattarai +4
For decades, corporations and governments have relied on scanned documents to record vast amounts of information. However, extracting this information is a slow and tedious process…