1 paper · 1 filter
Junjie Li, Ziao Wang, Jianghong Ma +1
Large vision-language models (VLMs) achieve strong benchmark performance, but controlling their behavior through instruction tuning remains difficult. Reducing the budget of instru…