60 citations · 109 across the 9 of their papers we have counts for
10 papers
Learning to Learn Better for Video Object Segmentation
Meng Lan, Jing Zhang, Lefei Zhang +1
Recently, the joint learning framework (JOINT) integrates matching based transductive reasoning and online inductive learning to achieve accurate and robust semi-supervised video o…
1st Workshop on Maritime Computer Vision (MaCVi) 2023: Challenge Results
Benjamin Kiefer, Matej Kristan, Janez Perš +70
The 1 Workshop on Maritime Computer Vision (MaCVi) 2023 focused on maritime computer vision for Unmanned Aerial Vehicles (UAV) and Unmanned Surface Vehicle (USV), and…
Rethinking Hierarchies in Pre-trained Plain Vision Transformer
Yufei Xu, Jing Zhang, Qiming Zhang +1
Self-supervised pre-training vision transformer (ViT) via masked image modeling (MIM) has been proven very effective. However, customized algorithms should be carefully designed fo…
RU-Net: Regularized Unrolling Network for Scene Graph Generation
Xin Lin, Changxing Ding, Jing Zhang +2
Scene graph generation (SGG) aims to detect objects and predict the relationships between each pair of objects. Existing SGG methods usually suffer from several issues, including 1…
DearKD: Data-Efficient Early Knowledge Distillation for Vision Transformers
Xianing Chen, Qiong Cao, Yujie Zhong +3
Transformers are successfully applied to computer vision due to their powerful modeling capacity with self-attention. However, the excellent performance of transformers heavily dep…
Learning Affordance Grounding from Exocentric Images
Hongchen Luo, Wei Zhai, Jing Zhang +2
Affordance grounding, a task to ground (i.e., localize) action possibility region in objects, which faces the challenge of establishing an explicit link with object parts due to th…