2 citations · 2 across the 2 of their papers we have counts for
2 papers
cs.CV2024
Vote&Mix: Plug-and-Play Token Reduction for Efficient Vision Transformer
Shuai Peng, Di Fu, Baole Wei +3
Despite the remarkable success of Vision Transformers (ViTs) in various visual tasks, they are often hindered by substantial computational cost. In this work, we introduce Vote\&Mi…
cs.CV2023★ 2 cited
Recognition-Guided Diffusion Model for Scene Text Image Super-Resolution
Yuxuan Zhou, Liangcai Gao, Zhi Tang +1
Scene Text Image Super-Resolution (STISR) aims to enhance the resolution and legibility of text within low-resolution (LR) images, consequently elevating recognition accuracy in Sc…