12 citations · 17 across the 4 of their papers we have counts for
4 papers
Attention in Attention: Modeling Context Correlation for Efficient Video Classification
Yanbin Hao, Shuo Wang, Pei Cao +4
Attention mechanisms have significantly boosted the performance of video classification neural networks thanks to the utilization of perspective contexts. However, the current rese…
Group Contextualization for Video Recognition
Yanbin Hao, Hao Zhang, Chong-Wah Ngo +1
Learning discriminative representation from the complex spatio-temporal dynamic space is essential for video recognition. On top of those stylized spatio-temporal computational uni…
Token Shift Transformer for Video Classification
Hao Zhang, Yanbin Hao, Chong-Wah Ngo
Transformer achieves remarkable successes in understanding 1 and 2-dimensional signals (e.g., NLP and Image Content Understanding). As a potential alternative to convolutional neur…
Aggregated Multi-GANs for Controlled 3D Human Motion Prediction
Zhenguang Liu, Kedi Lyu, Shuang Wu +3
Human motion prediction from historical pose sequence is at the core of many applications in machine intelligence. However, in current state-of-the-art methods, the predicted futur…