213 citations · 213 across the 1 of their papers we have counts for
1 paper
Jong Hak Moon, Hyungyung Lee, Woncheol Shin +2
Recently a number of studies demonstrated impressive performance on diverse vision-language multi-modal tasks such as image captioning and visual question answering by extending th…