3 papers
cs.CL2024
JaPOC: Japanese Post-OCR Correction Benchmark using Vouchers
Masato Fujitake
In this paper, we create benchmarks and assess the effectiveness of error correction methods for Japanese vouchers in OCR (Optical Character Recognition) systems. It is essential f…
cs.CV2024
JSTR: Judgment Improves Scene Text Recognition
Masato Fujitake
In this paper, we present a method for enhancing the accuracy of scene text recognition tasks by judging whether the image and text match each other. While previous studies focused…
cs.CL2024
LayoutLLM: Large Language Model Instruction Tuning for Visually Rich Document Understanding
Masato Fujitake
This paper proposes LayoutLLM, a more flexible document analysis method for understanding imaged documents. Visually Rich Document Understanding tasks, such as document image class…