207 citations · 988 across the 24 of their papers we have counts for
15 papers · 2 filters
Siamese Network with Interactive Transformer for Video Object Segmentation
Meng Lan, Jing Zhang, Fengxiang He +1
Semi-supervised video object segmentation (VOS) refers to segmenting the target object in remaining frames given its annotation in the first frame, which has been actively studied…
Recurrent Glimpse-based Decoder for Detection with Transformer
Zhe Chen, Jing Zhang, Dacheng Tao
Although detection with Transformer (DETR) is increasingly popular, its global attention modeling requires an extremely long training period to optimize and achieve promising detec…
PolyphonicFormer: Unified Query Learning for Depth-aware Video Panoptic Segmentation
Haobo Yuan, Xiangtai Li, Yibo Yang +5
The Depth-aware Video Panoptic Segmentation (DVPS) is a new challenging vision problem that aims to predict panoptic segmentation and depth in a video simultaneously. The previous…
FIBA: Frequency-Injection based Backdoor Attack in Medical Image Analysis
Yu Feng, Benteng Ma, Jing Zhang +3
In recent years, the security of AI systems has drawn increasing research attention, especially in the medical imaging realm. To develop a secure medical image analysis (MIA) syste…
GMFlow: Learning Optical Flow via Global Matching
Haofei Xu, Jing Zhang, Jianfei Cai +2
Learning-based optical flow estimation has been dominated with the pipeline of cost volume with convolutions for flow regression, which is inherently limited to local correlations…
RegionCL: Can Simple Region Swapping Contribute to Contrastive Learning?
Yufei Xu, Qiming Zhang, Jing Zhang +1
Self-supervised methods (SSL) have achieved significant success via maximizing the mutual information between two augmented views, where cropping is a popular augmentation techniqu…