activity
20162024
most citedMAMBA: Multi-level Aggregation via Memory Bank for Video Object Detection

60 citations · 80 across the 5 of their papers we have counts for

collaborators

5 papers

cs.CV20242 cited

TDViT: Temporal Dilated Video Transformer for Dense Video Tasks

Guanxiong Sun, Yang Hua, Guosheng Hu +1

Deep video models, for example, 3D CNNs or video transformers, have achieved promising performance on sparse video tasks, i.e., predicting one result per video. However, challenges…

cs.CV202417 cited

Efficient One-stage Video Object Detection by Exploiting Temporal Consistency

Guanxiong Sun, Yang Hua, Guosheng Hu +1

Recently, one-stage detectors have achieved competitive accuracy and faster speed compared with traditional two-stage detectors on image data. However, in the field of video object…

cs.CV202460 cited

MAMBA: Multi-level Aggregation via Memory Bank for Video Object Detection

Guanxiong Sun, Yang Hua, Guosheng Hu +1

State-of-the-art video object detection methods maintain a memory structure, either a sliding window or a memory queue, to enhance the current frame using attention mechanisms. How…

cs.LG20221 cited

ProSelfLC: Progressive Self Label Correction Towards A Low-Temperature Entropy State

Xinshao Wang, Yang Hua, Elyor Kodirov +3

There is a family of label modification approaches including self and non-self label correction (LC), and output regularisation. They are widely used for training robust deep neura…

cs.CV2016

Deep Convolutional Poses for Human Interaction Recognition in Monocular Videos

Marcel Sheeny de Moraes, Sankha Mukherjee, Neil M Robertson

Human interaction recognition is a challenging problem in computer vision and has been researched over the years due to its important applications. With the development of deep mod…