1 citations · 1 across the 3 of their papers we have counts for
1 paper · 1 filter
Riley Tavassoli, Mani Amani, Reza Akhavian
Vision-language models (VLMs) have shown powerful capabilities in visual question answering and reasoning tasks by combining visual representations with the abstract skill set larg…