10 citations · 10 across the 1 of their papers we have counts for
1 paper
Yiqing Huang, Jiansheng Chen
Existing image captioning models are usually trained by cross-entropy (XE) loss and reinforcement learning (RL), which set ground-truth words as hard targets and force the captioni…