203 citations · 265 across the 9 of their papers we have counts for
13 papers · 1 filter
End-to-End Human Instance Matting
Qinglin Liu, Shengping Zhang, Quanling Meng +3
Human instance matting aims to estimate an alpha matte for each human instance in an image, which is extremely challenging and has rarely been studied so far. Despite some efforts…
Towards Unified Token Learning for Vision-Language Tracking
Yaozong Zheng, Bineng Zhong, Qihua Liang +3
In this paper, we present a simple, flexible and effective vision-language (VL) tracking pipeline, termed \textbf{MMTrack}, which casts VL tracking as a token generation task. Trad…
Visualizing and Understanding Patch Interactions in Vision Transformer
Jie Ma, Yalong Bai, Bineng Zhong +3
Vision Transformer (ViT) has become a leading tool in various computer vision tasks, owing to its unique self-attention mechanism that learns visual representations explicitly thro…
Distractor-Aware Fast Tracking via Dynamic Convolutions and MOT Philosophy
Zikai Zhang, Bineng Zhong, Shengping Zhang +3
A practical long-term tracker typically contains three key properties, i.e. an efficient model design, an effective global re-detection strategy and a robust distractor awareness m…
Learning to Filter: Siamese Relation Network for Robust Tracking
Siyuan Cheng, Bineng Zhong, Guorong Li +4
Despite the great success of Siamese-based trackers, their performance under complicated scenarios is still not satisfying, especially when there are distractors. To this end, we p…
Projection & Probability-Driven Black-Box Attack
Jie Li, Rongrong Ji, Hong Liu +4
Generating adversarial examples in a black-box setting retains a significant challenge with vast practical application prospects. In particular, existing black-box attacks suffer f…