3 citations · 4 across the 4 of their papers we have counts for
1 paper · 1 filter
Jiawei Liang, Siyuan Liang, Man Luo +4
Autoregressive Visual Language Models (VLMs) showcase impressive few-shot learning capabilities in a multimodal context. Recently, multimodal instruction tuning has been proposed t…