60 citations · 89 across the 12 of their papers we have counts for
1 paper · 1 filter
Ugur Sahin, Hang Li, Qadeer Khan +2
Contemporary large-scale visual language models (VLMs) exhibit strong representation capacities, making them ubiquitous for enhancing image and text understanding tasks. They are o…