32 citations · 37 across the 5 of their papers we have counts for
7 papers
Unsupervised Pre-training for Temporal Action Localization Tasks
Can Zhang, Tianyu Yang, Junwu Weng +3
Unsupervised video representation learning has made remarkable achievements in recent years. However, most existing methods are designed and optimized for video classification. The…
Long-Short Temporal Modeling for Efficient Action Recognition
Liyu Wu, Yuexian Zou, Can Zhang
Efficient long-short temporal modeling is key for enhancing the performance of action recognition task. In this paper, we propose a new two-stream action recognition network, terme…
SRF-Net: Selective Receptive Field Network for Anchor-Free Temporal Action Detection
Ranyu Ning, Can Zhang, Yuexian Zou
Temporal action detection (TAD) is a challenging task which aims to temporally localize and recognize the human action in untrimmed videos. Current mainstream one-stage TAD approac…
All You Need is a Second Look: Towards Arbitrary-Shaped Text Detection
Meng Cao, Can Zhang, Dongming Yang +1
Arbitrary-shaped text detection is a challenging task since curved texts in the wild are of the complex geometric layouts. Existing mainstream methods follow the instance segmentat…
RR-Net: Injecting Interactive Semantics in Human-Object Interaction Detection
Dongming Yang, Yuexian Zou, Can Zhang +2
Human-Object Interaction (HOI) detection devotes to learn how humans interact with surrounding objects. Latest end-to-end HOI detectors are short of relation reasoning, which leads…
CoLA: Weakly-Supervised Temporal Action Localization with Snippet Contrastive Learning
Can Zhang, Meng Cao, Dongming Yang +2
Weakly-supervised temporal action localization (WS-TAL) aims to localize actions in untrimmed videos with only video-level labels. Most existing models follow the "localization by…