60 citations · 80 across the 5 of their papers we have counts for
5 papers
TDViT: Temporal Dilated Video Transformer for Dense Video Tasks
Guanxiong Sun, Yang Hua, Guosheng Hu +1
Deep video models, for example, 3D CNNs or video transformers, have achieved promising performance on sparse video tasks, i.e., predicting one result per video. However, challenges…
Efficient One-stage Video Object Detection by Exploiting Temporal Consistency
Guanxiong Sun, Yang Hua, Guosheng Hu +1
Recently, one-stage detectors have achieved competitive accuracy and faster speed compared with traditional two-stage detectors on image data. However, in the field of video object…
MAMBA: Multi-level Aggregation via Memory Bank for Video Object Detection
Guanxiong Sun, Yang Hua, Guosheng Hu +1
State-of-the-art video object detection methods maintain a memory structure, either a sliding window or a memory queue, to enhance the current frame using attention mechanisms. How…
ProSelfLC: Progressive Self Label Correction Towards A Low-Temperature Entropy State
Xinshao Wang, Yang Hua, Elyor Kodirov +3
There is a family of label modification approaches including self and non-self label correction (LC), and output regularisation. They are widely used for training robust deep neura…
Deep Convolutional Poses for Human Interaction Recognition in Monocular Videos
Marcel Sheeny de Moraes, Sankha Mukherjee, Neil M Robertson
Human interaction recognition is a challenging problem in computer vision and has been researched over the years due to its important applications. With the development of deep mod…