1 citations · 1 across the 3 of their papers we have counts for
4 papers
Platypus: A Generalized Specialist Model for Reading Text in Various Forms
Peng Wang, Zhaohai Li, Jun Tang +4
Reading text from images (either natural scenes or documents) has been a long-standing research topic for decades, due to the high technical challenge and wide application range. P…
Visual Text Generation in the Wild
Yuanzhi Zhu, Jiawei Liu, Feiyu Gao +6
Recently, with the rapid advancements of generative models, the field of visual text generation has witnessed significant progress. However, it is still challenging to render high-…
OmniParser: A Unified Framework for Text Spotting, Key Information Extraction and Table Recognition
Jianqiang Wan, Sibo Song, Wenwen Yu +6
Recently, visually-situated text parsing (VsTP) has experienced notable advancements, driven by the increasing demand for automated document understanding and the emergence of Gene…
LORE++: Logical Location Regression Network for Table Structure Recognition with Pre-training
Rujiao Long, Hangdi Xing, Zhibo Yang +4
Table structure recognition (TSR) aims at extracting tables in images into machine-understandable formats. Recent methods solve this problem by predicting the adjacency relations o…