48 citations · 78 across the 4 of their papers we have counts for
4 papers
Knowledge Integration Networks for Action Recognition
Shiwen Zhang, Sheng Guo, Limin Wang +2
In this work, we propose Knowledge Integration Networks (referred as KINet) for video action recognition. KINet is capable of aggregating meaningful context features which are of g…
V4D:4D Convolutional Neural Networks for Video-level Representation Learning
Shiwen Zhang, Sheng Guo, Weilin Huang +2
Most existing 3D CNNs for video representation learning are clip-based methods, and thus do not consider video-level temporal evolution of spatio-temporal features. In this paper,…
Convolutional Character Networks
Linjie Xing, Zhi Tian, Weilin Huang +1
Recent progress has been made on developing a unified framework for joint text detection and recognition in natural images, but existing joint models were mostly built on two-stage…
Compatible and Diverse Fashion Image Inpainting
Xintong Han, Zuxuan Wu, Weilin Huang +2
Visual compatibility is critical for fashion analysis, yet is missing in existing fashion image synthesis systems. In this paper, we propose to explicitly model visual compatibilit…