9 citations · 9 across the 2 of their papers we have counts for
2 papers
cs.CV2023
Token Sparsification for Faster Medical Image Segmentation
Lei Zhou, Huidong Liu, Joseph Bae +3
Can we use sparse tokens for dense prediction, e.g., segmentation? Although token sparsification has been applied to Vision Transformers (ViT) to accelerate classification, it is s…
cs.CV2021★ 9 cited
CMA-CLIP: Cross-Modality Attention CLIP for Image-Text Classification
Huidong Liu, Shaoyuan Xu, Jinmiao Fu +5
Modern Web systems such as social media and e-commerce contain rich contents expressed in images and text. Leveraging information from multi-modalities can improve the performance…