3 citations · 3 across the 1 of their papers we have counts for
1 paper
Deepanway Ghosal, Vernon Toh Yan Han, Chia Yew Ken +1
This paper introduces the novel task of multimodal puzzle solving, framed within the context of visual question-answering. We present a new dataset, AlgoPuzzleVQA designed to chall…