20 citations · 20 across the 1 of their papers we have counts for
3 papers
cs.CV2023
GlobalDoc: A Cross-Modal Vision-Language Framework for Real-World Document Image Retrieval and Classification
Souhail Bakkali, Sanket Biswas, Zuheng Ming +4
Visual document understanding (VDU) has rapidly advanced with the development of powerful multi-modal language models. However, these models typically require extensive document pr…
cs.CV2020★ 20 cited
Pay Attention to What You Read: Non-recurrent Handwritten Text-Line Recognition
Lei Kang, Pau Riba, Marçal Rusiñol +2
The advent of recurrent neural networks for handwriting recognition marked an important milestone reaching impressive recognition accuracies despite the great variability that we o…
cs.CV2019
Selective Style Transfer for Text
Raul Gomez, Ali Furkan Biten, Lluis Gomez +3
This paper explores the possibilities of image style transfer applied to text maintaining the original transcriptions. Results on different text domains (scene text, machine printe…