2 citations · 2 across the 3 of their papers we have counts for
1 paper · 1 filter
Pete Janowczyk, Linda Laurier, Ave Giulietta +2
Multi-Modal Language Models (MLLMs) have transformed artificial intelligence by combining visual and text data, making applications like image captioning, visual question answering…