1 paper
Jihae Jeong, Junha Choi, Hwanjo Yu
Large vision-language models (LVLMs) often hallucinate, generating content that the input image does not support. Preventing such content during decoding calls for a candidate-spec…