2 citations · 3 across the 3 of their papers we have counts for
3 papers
cs.CL2024★ 1 cited
MultiMath: Bridging Visual and Mathematical Reasoning for Large Language Models
Shuai Peng, Di Fu, Liangcai Gao +3
The rapid development of large language models (LLMs) has spurred extensive research into their domain-specific capabilities, particularly mathematical reasoning. However, most ope…
cs.CV2024
Vote&Mix: Plug-and-Play Token Reduction for Efficient Vision Transformer
Shuai Peng, Di Fu, Baole Wei +3
Despite the remarkable success of Vision Transformers (ViTs) in various visual tasks, they are often hindered by substantial computational cost. In this work, we introduce Vote\&Mi…
cs.CV2023★ 2 cited
Recognition-Guided Diffusion Model for Scene Text Image Super-Resolution
Yuxuan Zhou, Liangcai Gao, Zhi Tang +1
Scene Text Image Super-Resolution (STISR) aims to enhance the resolution and legibility of text within low-resolution (LR) images, consequently elevating recognition accuracy in Sc…