57 citations · 60 across the 11 of their papers we have counts for
5 papers · 1 filter
A Simple yet Effective Layout Token in Large Language Models for Document Understanding
Zhaoqing Zhu, Chuwei Luo, Zirui Shao +4
Recent methods that integrate spatial layouts with text for document understanding in large language models (LLMs) have shown promising results. A commonly used method is to repres…
WebRPG: Automatic Web Rendering Parameters Generation for Visual Presentation
Zirui Shao, Feiyu Gao, Hangdi Xing +5
In the era of content creation revolution propelled by advancements in generative models, the field of web design remains unexplored despite its critical role in modern digital com…
Visual Text Generation in the Wild
Yuanzhi Zhu, Jiawei Liu, Feiyu Gao +6
Recently, with the rapid advancements of generative models, the field of visual text generation has witnessed significant progress. However, it is still challenging to render high-…
LORE: Logical Location Regression Network for Table Structure Recognition
Hangdi Xing, Feiyu Gao, Rujiao Long +5
Table structure recognition (TSR) aims at extracting tables in images into machine-understandable formats. Recent methods solve this problem by predicting the adjacency relations o…
Parsing Table Structures in the Wild
Rujiao Long, Wen Wang, Nan Xue +4
This paper tackles the problem of table structure parsing (TSP) from images in the wild. In contrast to existing studies that mainly focus on parsing well-aligned tabular images wi…