2 citations · 2 across the 1 of their papers we have counts for
1 paper
Wenqiao Zhang, Haochen Shi, Siliang Tang +3
The contemporary visual captioning models frequently hallucinate objects that are not actually in a scene, due to the visual misclassification or over-reliance on priors that resul…