3 citations · 3 across the 2 of their papers we have counts for
Showing cs.CVShow all
3 papers · 1 filter
cs.CV2023★ 3 cited
DTrOCR: Decoder-only Transformer for Optical Character Recognition
Masato Fujitake
Typical text recognition methods rely on an encoder-decoder structure, in which the encoder extracts features from an image, and the decoder produces recognized text from these fea…
cs.CV2023
DiffusionSTR: Diffusion Model for Scene Text Recognition
Masato Fujitake
This paper presents Diffusion Model for Scene Text Recognition (DiffusionSTR), an end-to-end text recognition framework using diffusion models for recognizing text in the wild. Whi…
cs.CV2023
A3S: Adversarial learning of semantic representations for Scene-Text Spotting
Masato Fujitake
Scene-text spotting is a task that predicts a text area on natural scene images and recognizes its text characters simultaneously. It has attracted much attention in recent years d…