3 citations · 3 across the 4 of their papers we have counts for
5 papers
1st Place Solution to ECCV 2022 Challenge on Out of Vocabulary Scene Text Understanding: End-to-End Recognition of Out of Vocabulary Words
Zhangzi Zhu, Chuhui Xue, Yu Hao +2
Scene text recognition has attracted increasing interest in recent years due to its wide range of applications in multilingual translation, autonomous driving, etc. In this report,…
Runner-Up Solution to ECCV 2022 Challenge on Out of Vocabulary Scene Text Understanding: Cropped Word Recognition
Zhangzi Zhu, Yu Hao, Wenqing Zhang +2
This report presents our 2nd place solution to ECCV 2022 challenge on Out-of-Vocabulary Scene Text Understanding (OOV-ST) : Cropped Word Recognition. This challenge is held in the…
Improving Image Captioning with Control Signal of Sentence Quality
Zhangzi Zhu, Hong Qu
In the dataset of image captioning, each image is aligned with several descriptions. Despite the fact that the quality of these descriptions varies, existing captioning models trea…
Self-Annotated Training for Controllable Image Captioning
Zhangzi Zhu, Tianlei Wang, Hong Qu
The Controllable Image Captioning (CIC) task aims to generate captions conditioned on designated control signals. Several structure-related control signals are proposed to control…
Macroscopic Control of Text Generation for Image Captioning
Zhangzi Zhu, Tianlei Wang, Hong Qu
Despite the fact that image captioning models have been able to generate impressive descriptions for a given image, challenges remain: (1) the controllability and diversity of exis…