1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.CV2024
CountFormer: Multi-View Crowd Counting Transformer
Hong Mo, Xiong Zhang, Jianchao Tan +4
Multi-view counting (MVC) methods have shown their superiority over single-view counterparts, particularly in situations characterized by heavy occlusion and severe perspective dis…
cs.CV2022★ 1 cited
Convolutional Embedding Makes Hierarchical Vision Transformer Stronger
Cong Wang, Hongmin Xu, Xiong Zhang +3
Vision Transformers (ViTs) have recently dominated a range of computer vision tasks, yet it suffers from low training data efficiency and inferior local semantic representation cap…