1 citations · 1 across the 14 of their papers we have counts for
1 paper · 1 filter
Wen Wen, Tianwu Zhi, Kanglong Fan +6
Improving vision-language models (VLMs) in the post-training stage typically relies on supervised fine-tuning or reinforcement learning, methods that necessitate costly, human-anno…