2 citations · 2 across the 4 of their papers we have counts for
1 paper · 1 filter
Dimitrios Damianos, Leon Voukoutis, Georgios Skyrianos +2
Generative Vision-Language Models (VLMs) perform well on multimodal reasoning, but how visual inputs are transformed to text remains poorly understood. Existing interpretability wo…