13 citations · 13 across the 2 of their papers we have counts for
2 papers
cs.CV2025
Compress image to patches for Vision Transformer
Xinfeng Zhao, Yaoru Sun
The Vision Transformer (ViT) has made significant strides in the field of computer vision. However, as the depth of the model and the resolution of the input images increase, the c…
cs.CV2018★ 13 cited
Cross-domain Human Parsing via Adversarial Feature and Label Adaptation
Si Liu, Yao Sun, Defa Zhu +4
Human parsing has been extensively studied recently due to its wide applications in many important scenarios. Mainstream fashion parsing models focus on parsing the high-resolution…