1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.CV2022
End-to-end video instance segmentation via spatial-temporal graph neural networks
Tao Wang, Ning Xu, Kean Chen +1
Video instance segmentation is a challenging task that extends image instance segmentation to the video domain. Existing methods either rely only on single-frame information for th…
cs.CV2022★ 1 cited
Uni-EDEN: Universal Encoder-Decoder Network by Multi-Granular Vision-Language Pre-training
Yehao Li, Jiahao Fan, Yingwei Pan +3
Vision-language pre-training has been an emerging and fast-developing research topic, which transfers multi-modal knowledge from rich-resource pre-training task to limited-resource…