2 citations · 2 across the 1 of their papers we have counts for
2 papers
cs.CV2021★ 2 cited
Semi-Autoregressive Transformer for Image Captioning
Yuanen Zhou, Yong Zhang, Zhenzhen Hu +1
Current state-of-the-art image captioning models adopt autoregressive decoders, \ie they generate each word by conditioning on previously generated words, which leads to heavy late…
cs.CV2020
More Grounded Image Captioning by Distilling Image-Text Matching Model
Yuanen Zhou, Meng Wang, Daqing Liu +2
Visual attention not only improves the performance of image captioners, but also serves as a visual interpretation to qualitatively measure the caption rationality and model transp…