28 citations · 82 across the 14 of their papers we have counts for
21 papers · 1 filter
Task Grouping for Multilingual Text Recognition
Jing Huang, Kevin J Liang, Rama Kovvuri +1
Most existing OCR methods focus on alphanumeric characters due to the popularity of English and numbers, as well as their corresponding datasets. On extending the characters to mor…
Proactive Image Manipulation Detection
Vishal Asnani, Xi Yin, Tal Hassner +2
Image manipulation detection algorithms are often trained to discriminate between images manipulated with particular Generative Models (GMs) and genuine/real images, yet generalize…
FSGANv2: Improved Subject Agnostic Face Swapping and Reenactment
Yuval Nirkin, Yosi Keller, Tal Hassner
We present Face Swapping GAN (FSGAN) for face swapping and reenactment. Unlike previous work, we offer a subject agnostic swapping scheme that can be applied to pairs of faces with…
TextStyleBrush: Transfer of Text Aesthetics from a Single Example
Praveen Krishnan, Rama Kovvuri, Guan Pang +2
We present a novel approach for disentangling the content of a text image from all aspects of its appearance. The appearance representation we derive can then be applied to new con…
TextOCR: Towards large-scale end-to-end reasoning for arbitrary-shaped scene text
Amanpreet Singh, Guan Pang, Mandy Toh +3
A crucial component for the scene text based reasoning required for TextVQA and TextCaps datasets involve detecting and recognizing text present in the images using an optical char…
A Multiplexed Network for End-to-End, Multilingual OCR
Jing Huang, Guan Pang, Rama Kovvuri +5
Recent advances in OCR have shown that an end-to-end (E2E) training pipeline that includes both detection and recognition leads to the best results. However, many existing methods…