3 citations · 4 across the 3 of their papers we have counts for
3 papers
cs.CV2022
PPMN: Pixel-Phrase Matching Network for One-Stage Panoptic Narrative Grounding
Zihan Ding, Zi-han Ding, Tianrui Hui +4
Panoptic Narrative Grounding (PNG) is an emerging task whose goal is to segment visual objects of things and stuff categories described by dense narrative captions of a still image…
cs.CV2022★ 3 cited
MT-Net Submission to the Waymo 3D Detection Leaderboard
Shaoxiang Chen, Zequn Jie, Xiaolin Wei +1
In this technical report, we introduce our submission to the Waymo 3D Detection leaderboard. Our network is based on the Centerpoint architecture, but with significant improvements…
cs.CV2022★ 1 cited
Efficient Modeling of Future Context for Image Captioning
Zhengcong Fei, Junshi Huang, Xiaoming Wei +1
Existing approaches to image captioning usually generate the sentence word-by-word from left to right, with the constraint of conditioned on local context including the given image…