1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.CV2026
BabelDOC: Better Layout-Preserving PDF Translation via Intermediate Representation
Qi Yang, Xiangyao Ma, Xiao Wang +2
As global cross-lingual communication intensifies, language barriers in visually rich documents such as PDFs remain a practical bottleneck. Existing document translation pipelines…
cs.HC2023★ 1 cited
DocTrack: A Visually-Rich Document Dataset Really Aligned with Human Eye Movement for Machine Reading
Hao Wang, Qingxuan Wang, Yue Li +3
The use of visually-rich documents (VRDs) in various fields has created a demand for Document AI models that can read and comprehend documents like humans, which requires the overc…
cs.CV2023
Vision-Enhanced Semantic Entity Recognition in Document Images via Visually-Asymmetric Consistency Learning
Hao Wang, Xiahua Chen, Rui Wang +1
Extracting meaningful entities belonging to predefined categories from Visually-rich Form-like Documents (VFDs) is a challenging task. Visual and layout features such as font, back…