253 citations · 739 across the 29 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2023★ 6 cited
DSVT: Dynamic Sparse Voxel Transformer with Rotated Sets
Haiyang Wang, Chen Shi, Shaoshuai Shi +5
Designing an efficient yet deployment-friendly 3D backbone to handle sparse point clouds is a fundamental problem in 3D perception. Compared with the customized sparse convolution,…
cs.CV2021★ 8 cited
Can Vision Transformers Perform Convolution?
Shanda Li, Xiangning Chen, Di He +1
Several recent studies have demonstrated that attention-based networks, such as Vision Transformer (ViT), can outperform Convolutional Neural Networks (CNNs) on several computer vi…