6 citations · 22 across the 6 of their papers we have counts for
1 paper · 1 filter
Yekun Chai, Shuo Jin, Junliang Xing
Automatically translating images to texts involves image scene understanding and language modeling. In this paper, we propose a novel model, termed RefineCap, that refines the outp…