11 citations · 11 across the 1 of their papers we have counts for
1 paper
Hugo Laurençon, Andrés Marafioti, Victor Sanh +1
The field of vision-language models (VLMs), which take images and texts as inputs and output texts, is rapidly evolving and has yet to reach consensus on several key aspects of the…