24 citations · 32 across the 3 of their papers we have counts for
4 papers
MiniViT: Compressing Vision Transformers with Weight Multiplexing
Jinnian Zhang, Houwen Peng, Kan Wu +4
Vision Transformer (ViT) models have recently drawn much attention in computer vision due to their high model capability. However, ViT models suffer from huge number of parameters,…
Rethinking and Improving Relative Position Encoding for Vision Transformer
Kan Wu, Houwen Peng, Minghao Chen +2
Relative position encoding (RPE) is important for transformer to capture sequence ordering of input tokens. General efficacy has been proven in natural language processing. However…
LightTrack: Finding Lightweight Neural Networks for Object Tracking via One-Shot Architecture Search
Bin Yan, Houwen Peng, Kan Wu +3
Object tracking has achieved significant progress over the past few years. However, state-of-the-art trackers become increasingly heavy and expensive, which limits their deployment…
Harvesting Visual Objects from Internet Images via Deep Learning Based Objectness Assessment
Kan Wu, Guanbin Li, Haofeng Li +2
The collection of internet images has been growing in an astonishing speed. It is undoubted that these images contain rich visual information that can be useful in many application…