11 citations · 15 across the 3 of their papers we have counts for
4 papers
Blending Anti-Aliasing into Vision Transformer
Shengju Qian, Hao Shao, Yi Zhu +2
The transformer architectures, based on self-attention mechanism and convolution-free design, recently found superior performance and booming applications in computer vision. Howev…
1st place solution for AVA-Kinetics Crossover in AcitivityNet Challenge 2020
Siyu Chen, Junting Pan, Guanglu Song +6
This technical report introduces our winning solution to the spatio-temporal action localization track, AVA-Kinetics Crossover, in ActivityNet Challenge 2020. Our entry is mainly b…
Top-1 Solution of Multi-Moments in Time Challenge 2019
Manyuan Zhang, Hao Shao, Guanglu Song +2
In this technical report, we briefly introduce the solutions of our team 'Efficient' for the Multi-Moments in Time challenge in ICCV 2019. We first conduct several experiments with…
Temporal Interlacing Network
Hao Shao, Shengju Qian, Yu Liu
For a long time, the vision community tries to learn the spatio-temporal representation by combining convolutional neural network together with various temporal models, such as the…