18 citations · 18 across the 1 of their papers we have counts for
2 papers
cs.CV2021
Dual-Level Collaborative Transformer for Image Captioning
Yunpeng Luo, Jiayi Ji, Xiaoshuai Sun +5
Descriptive region features extracted by object detection networks have played an important role in the recent advancements of image captioning. However, they are still criticized…
cs.CV2020★ 18 cited
Improving Image Captioning by Leveraging Intra- and Inter-layer Global Representation in Transformer Network
Jiayi Ji, Yunpeng Luo, Xiaoshuai Sun +5
Transformer-based architectures have shown great success in image captioning, where object regions are encoded and then attended into the vectorial representations to guide the cap…