1 paper · 1 filter
Chenglong Wang, Yang Gan, Yifu Huo +9
Large vision-language models (LVLMs) often fail to align with human preferences, leading to issues like generating misleading content without proper visual context (also known as h…