2 citations · 3 across the 4 of their papers we have counts for
4 papers
Reading Order Matters: Information Extraction from Visually-rich Documents by Token Path Prediction
Chong Zhang, Ya Guo, Yi Tu +5
Recent advances in multimodal pre-trained models have significantly improved information extraction from visually-rich documents (VrDs), in which named entity recognition (NER) is…
Teukolsky-like equations with various spins in spherically symmetric spacetime
Ya Guo, Hiroaki Nakajima, Wenbin Lin
We study the wave equations with various spins on the background of the general spherically symmetric spacetime. We obtain the unified expression of the Teukolsky-like master equat…
LayoutMask: Enhance Text-Layout Interaction in Multi-modal Pre-training for Document Understanding
Yi Tu, Ya Guo, Huan Chen +1
Visually-rich Document Understanding (VrDU) has attracted much research attention over the past years. Pre-trained models on a large number of document images with transformer-base…
Gravitational-wave equation in effective one-body background for spinless binary
Ya Guo, Hiroaki Nakajima, Wenbin Lin
We construct the gravitational-wave equation in the background of the effective one-body system for the spinless binary, which is in general available with the spherically symmetri…