1 paper
Zijie Song, Yuanen Zhou, Zhenzhen Hu +4
Most current image captioning models typically generate captions from left-to-right. This unidirectional property makes them can only leverage past context but not future context.…