3 citations · 5 across the 2 of their papers we have counts for
1 paper · 1 filter
Zhiyang Xu, Chao Feng, Rulin Shao +6
Despite vision-language models' (VLMs) remarkable capabilities as versatile visual assistants, two substantial challenges persist within the existing VLM frameworks: (1) lacking ta…