1 paper
Yijin Ni, Simon Yu, Peng Qi
Vision and language models frequently ignore semantically critical input edits, defaulting to pretraining priors. For example, models will confidently assert a five-legged dog has…