3 citations · 3 across the 2 of their papers we have counts for
3 papers
cs.CV2024
Effectively Enhancing Vision Language Large Models by Prompt Augmentation and Caption Utilization
Minyi Zhao, Jie Wang, Zhaoyang Li +3
Recent studies have shown that Vision Language Large Models (VLLMs) may output content not relevant to the input images. This problem, called the hallucination phenomenon, undoubte…
cs.CV2023
HiREN: Towards Higher Supervision Quality for Better Scene Text Image Super-Resolution
Minyi Zhao, Yi Xu, Bingjia Li +3
Scene text image super-resolution (STISR) is an important pre-processing technique for text recognition from low-resolution scene images. Nowadays, various methods have been propos…
cs.CV2022★ 3 cited
C3-STISR: Scene Text Image Super-resolution with Triple Clues
Minyi Zhao, Miao Wang, Fan Bai +3
Scene text image super-resolution (STISR) has been regarded as an important pre-processing task for text recognition from low-resolution scene text images. Most recent approaches u…