6 citations · 6 across the 2 of their papers we have counts for
2 papers
cs.CV2024
CREPE: Coordinate-Aware End-to-End Document Parser
Yamato Okamoto, Youngmin Baek, Geewook Kim +5
In this study, we formulate an OCR-free sequence generation model for visual document understanding (VDU). Our model not only parses text from document images but also extracts the…
cs.CV2022★ 6 cited
DEER: Detection-agnostic End-to-End Recognizer for Scene Text Spotting
Seonghyeon Kim, Seung Shin, Yoonsik Kim +6
Recent end-to-end scene text spotters have achieved great improvement in recognizing arbitrary-shaped text instances. Common approaches for text spotting use region of interest poo…