2 citations · 4 across the 2 of their papers we have counts for
1 paper · 1 filter
Yang Ding, Jing Yu, Bang Liu +3
Knowledge-based visual question answering requires the ability of associating external knowledge for open-ended cross-modal scene understanding. One limitation of existing solution…