6 citations · 6 across the 3 of their papers we have counts for
3 papers · 1 filter
Kandinsky: an Improved Text-to-Image Synthesis with Image Prior and Latent Diffusion
Anton Razzhigaev, Arseniy Shakhmatov, Anastasia Maltseva +7
Text-to-image generation is a significant domain in modern computer vision and has achieved substantial improvements through the evolution of generative architectures. Among these,…
RusTitW: Russian Language Text Dataset for Visual Text in-the-Wild Recognition
Igor Markov, Sergey Nesteruk, Andrey Kuznetsov +1
Information surrounds people in modern life. Text is a very efficient type of information that people use for communication for centuries. However, automated text-in-the-wild recog…
Handwritten text generation and strikethrough characters augmentation
Alex Shonenkov, Denis Karachev, Max Novopoltsev +3
We introduce two data augmentation techniques, which, used with a Resnet-BiLSTM-CTC network, significantly reduce Word Error Rate (WER) and Character Error Rate (CER) beyond best-r…