29 citations · 37 across the 3 of their papers we have counts for
1 paper · 1 filter
Zheng Yuan, Qiao Jin, Chuanqi Tan +4
Vision-and-language multi-modal pretraining and fine-tuning have shown great success in visual question answering (VQA). Compared to general domain VQA, the performance of biomedic…