3 citations · 3 across the 1 of their papers we have counts for
5 papers · 1 filter
EfficientOCR: An Extensible, Open-Source Package for Efficiently Digitizing World Knowledge
Tom Bryan, Jacob Carlson, Abhishek Arora +1
Billions of public domain documents remain trapped in hard copy or lack an accurate digitization. Modern natural language processing methods cannot be used to index, retrieve, and…
Linking Representations with Multimodal Contrastive Learning
Abhishek Arora, Xinmei Yang, Shao-Yu Jheng +1
Many applications require linking individuals, firms, or locations across datasets. Most widely used methods, especially in social science, do not employ deep learning, with record…
Efficient OCR for Building a Diverse Digital History
Jacob Carlson, Tom Bryan, Melissa Dell
Thousands of users consult digital archives daily, but the information they can access is unrepresentative of the diversity of documentary history. The sequence-to-sequence archite…
LayoutParser: A Unified Toolkit for Deep Learning Based Document Image Analysis
Zejiang Shen, Ruochen Zhang, Melissa Dell +3
Recent advances in document image analysis (DIA) have been primarily driven by the application of neural networks. Ideally, research outcomes could be easily deployed in production…
A Large Dataset of Historical Japanese Documents with Complex Layouts
Zejiang Shen, Kaixuan Zhang, Melissa Dell
Deep learning-based approaches for automatic document layout analysis and content extraction have the potential to unlock rich information trapped in historical documents on a larg…