activity
20122022
most citedChannel-wise Topology Refinement Graph Convolution for Skeleton-Based Action Recognition

52 citations · 135 across the 12 of their papers we have counts for

collaborators
Showing cs.CVShow all

13 papers · 1 filter

cs.CV20225 cited

CREATE: A Benchmark for Chinese Short Video Retrieval and Title Generation

Ziqi Zhang, Yuxin Chen, Zongyang Ma +5

Previous works of video captioning aim to objectively describe the video's actual content, which lacks subjective and attractive expression, limiting its practical application scen…

cs.CV20223 cited

Learning Target-aware Representation for Visual Tracking via Informative Interactions

Mingzhe Guo, Zhipeng Zhang, Heng Fan +4

We introduce a novel backbone architecture to improve target-perception ability of feature representation for tracking. Specifically, having observed that de facto frameworks perfo…

cs.CV2021

SDTP: Semantic-aware Decoupled Transformer Pyramid for Dense Image Prediction

Zekun Li, Yufan Liu, Bing Li +3

Although transformer has achieved great progress on computer vision tasks, the scale variation in dense image prediction is still the key challenge. Few effective multi-scale techn…

cs.CV202152 cited

Channel-wise Topology Refinement Graph Convolution for Skeleton-Based Action Recognition

Yuxin Chen, Ziqi Zhang, Chunfeng Yuan +3

Graph convolutional networks (GCNs) have been widely used and achieved remarkable results in skeleton-based action recognition. In GCNs, graph topology dominates feature aggregatio…

cs.CV202112 cited

Learn to Match: Automatic Matching Network Design for Visual Tracking

Zhipeng Zhang, Yihao Liu, Xiao Wang +2

Siamese tracking has achieved groundbreaking performance in recent years, where the essence is the efficient matching operator cross-correlation and its variants. Besides the remar…

cs.CV2021

A Simple and Strong Baseline for Universal Targeted Attacks on Siamese Visual Tracking

Zhenbang Li, Yaya Shi, Jin Gao +4

Siamese trackers are shown to be vulnerable to adversarial attacks recently. However, the existing attack methods craft the perturbations for each video independently, which comes…