109 citations · 120 across the 2 of their papers we have counts for
2 papers
cs.CV2021★ 109 cited
CPTR: Full Transformer Network for Image Captioning
Wei Liu, Sihan Chen, Longteng Guo +2
In this paper, we consider the image captioning task from a new sequence-to-sequence prediction perspective and propose CaPtion TransformeR (CPTR) which takes the sequentialized ra…
cs.CV2021★ 11 cited
Global-Local Propagation Network for RGB-D Semantic Segmentation
Sihan Chen, Xinxin Zhu, Wei Liu +2
Depth information matters in RGB-D semantic segmentation task for providing additional geometric information to color images. Most existing methods exploit a multi-stage fusion str…