1 paper
Zhaokai Wang, Renda Bao, Qi Wu +1
When describing an image, reading text in the visual scene is crucial to understand the key information. Recent work explores the TextCaps task, i.e. image captioning with reading…