1 paper
Pierre Marza, Corentin Kervadec, Grigory Antipov +2
As in many tasks combining vision and language, both modalities play a crucial role in Visual Question Answering (VQA). To properly solve the task, a given model should both unders…