1 paper
Ofir Azachi, Kfir Eliyahu, Eyal El Ani +4
Hallucinations of vision-language models (VLMs), which are misalignments between visual content and generated text, undermine the reliability of VLMs. One common approach for detec…