1 paper
Viktor Kewenig, Andrew Lampinen, Samuel A. Nastase +5
Humans routinely draw on visual context to predict upcoming words. To what extent current vision-language models produce comparable behaviour is unclear. Here we placed five state-…