48 citations · 56 across the 4 of their papers we have counts for
4 papers · 1 filter
Forgedit: Text Guided Image Editing via Learning and Forgetting
Shiwen Zhang, Shuai Xiao, Weilin Huang
Text-guided image editing on real or synthetic images, given only the original image itself and the target text prompt as inputs, is a very general and challenging task. It require…
TFCNet: Temporal Fully Connected Networks for Static Unbiased Temporal Reasoning
Shiwen Zhang
Temporal Reasoning is one important functionality for vision intelligence. In computer vision research community, temporal reasoning is usually studied in the form of video classif…
Knowledge Integration Networks for Action Recognition
Shiwen Zhang, Sheng Guo, Limin Wang +2
In this work, we propose Knowledge Integration Networks (referred as KINet) for video action recognition. KINet is capable of aggregating meaningful context features which are of g…
V4D:4D Convolutional Neural Networks for Video-level Representation Learning
Shiwen Zhang, Sheng Guo, Weilin Huang +2
Most existing 3D CNNs for video representation learning are clip-based methods, and thus do not consider video-level temporal evolution of spatio-temporal features. In this paper,…