21 citations · 22 across the 3 of their papers we have counts for
4 papers
Exploiting Spatial-Temporal Semantic Consistency for Video Scene Parsing
Xingjian He, Weining Wang, Zhiyong Xu +3
Compared with image scene parsing, video scene parsing introduces temporal information, which can effectively improve the consistency and accuracy of prediction. In this paper, we…
OPT: Omni-Perception Pre-Trainer for Cross-Modal Understanding and Generation
Jing Liu, Xinxin Zhu, Fei Liu +8
In this paper, we propose an Omni-perception Pre-Trainer (OPT) for cross-modal understanding and generation, by jointly modeling visual, text and audio resources. OPT is constructe…
Face Sketch Synthesis via Semantic-Driven Generative Adversarial Network
Xingqun Qi, Muyi Sun, Weining Wang +3
Face sketch synthesis has made significant progress with the development of deep neural networks in these years. The delicate depiction of sketch portraits facilitates a wide range…
Temporal Memory Attention for Video Semantic Segmentation
Hao Wang, Weining Wang, Jing Liu
Video semantic segmentation requires to utilize the complex temporal relations between frames of the video sequence. Previous works usually exploit accurate optical flow to leverag…