1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.CV2023★ 1 cited
ESTextSpotter: Towards Better Scene Text Spotting with Explicit Synergy in Transformer
Mingxin Huang, Jiaxin Zhang, Dezhi Peng +5
In recent years, end-to-end scene text spotting approaches are evolving to the Transformer-based framework. While previous studies have shown the crucial importance of the intrinsi…
cs.CL2022
Knowing Where and What: Unified Word Block Pretraining for Document Understanding
Song Tao, Zijian Wang, Tiantian Fan +2
Due to the complex layouts of documents, it is challenging to extract information for documents. Most previous studies develop multimodal pre-trained models in a self-supervised wa…