most citedCrossFormer: A Versatile Vision Transformer Hinging on Cross-scale Attention

85 citations · 127 across the 7 of their papers we have counts for

collaborators

11 papers

cs.CV20222 cited

Boosting Semi-Supervised 3D Object Detection with Semi-Sampling

Xiaopei Wu, Yang Zhao, Liang Peng +6

Current 3D object detection methods heavily rely on an enormous amount of annotations. Semi-supervised learning can be used to alleviate this issue. Previous semi-supervised 3D obj…

cs.CV20221 cited

CLRNet: Cross Layer Refinement Network for Lane Detection

Tu Zheng, Yifei Huang, Yang Liu +4

Lane is critical in the vision navigation system of the intelligent vehicle. Naturally, lane is a traffic sign with high-level semantics, whereas it owns the specific local pattern…

cs.CV202212 cited

WeakM3D: Towards Weakly Supervised Monocular 3D Object Detection

Liang Peng, Senbo Yan, Boxi Wu +3

Monocular 3D object detection is one of the most challenging tasks in 3D scene understanding. Due to the ill-posed nature of monocular imagery, existing monocular 3D detection meth…

cs.CV202185 cited

CrossFormer: A Versatile Vision Transformer Hinging on Cross-scale Attention

Wenxiao Wang, Lu Yao, Long Chen +4

Transformers have made great progress in dealing with computer vision tasks. However, existing vision transformers do not yet possess the ability of building the interactions among…

cs.CV20213 cited

Salient Object Ranking with Position-Preserved Attention

Hao Fang, Daoxin Zhang, Yi Zhang +5

Instance segmentation can detect where the objects are in an image, but hard to understand the relationship between them. We pay attention to a typical relationship, relative salie…

cs.LG202113 cited

Attacking Adversarial Attacks as A Defense

Boxi Wu, Heng Pan, Li Shen +6

It is well known that adversarial attacks can fool deep neural networks with imperceptible perturbations. Although adversarial training significantly improves model robustness, fai…