activity
20182021
most citedJoint Visual Semantic Reasoning: Multi-Stage Decoder for Text Recognition

4 citations · 7 across the 2 of their papers we have counts for

collaborators

6 papers

cs.CV20214 cited

Joint Visual Semantic Reasoning: Multi-Stage Decoder for Text Recognition

Ayan Kumar Bhunia, Aneeshan Sain, Amandeep Kumar +3

Although text recognition has significantly evolved over the years, state-of-the-art (SOTA) models still struggle in the wild scenarios due to complex backgrounds, varying fonts, u…

cs.CV20213 cited

MetaHTR: Towards Writer-Adaptive Handwritten Text Recognition

Ayan Kumar Bhunia, Shuvozit Ghose, Amandeep Kumar +3

Handwritten Text Recognition (HTR) remains a challenging problem to date, largely due to the varying writing styles that exist amongst us. Prior works however generally operate wit…

cs.CV2020

UDBNET: Unsupervised Document Binarization Network via Adversarial Game

Amandeep Kumar, Shuvozit Ghose, Pinaki Nath Chowdhury +2

Degraded document image binarization is one of the most challenging tasks in the domain of document image analysis. In this paper, we present a novel approach towards document imag…

cs.CV2020

Modeling Extent-of-Texture Information for Ground Terrain Recognition

Shuvozit Ghose, Pinaki Nath Chowdhury, Partha Pratim Roy +1

Ground Terrain Recognition is a difficult task as the context information varies significantly over the regions of a ground terrain image. In this paper, we propose a novel approac…

cs.CV2018

A Deep One-Shot Network for Query-based Logo Retrieval

Ayan Kumar Bhunia, Ankan Kumar Bhunia, Shuvozit Ghose +3

Logo detection in real-world scene images is an important problem with applications in advertisement and marketing. Existing general-purpose object detection methods require large…

cs.CV2018

User Constrained Thumbnail Generation using Adaptive Convolutions

Perla Sai Raj Kishore, Ayan Kumar Bhunia, Shuvozit Ghose +1

Thumbnails are widely used all over the world as a preview for digital images. In this work we propose a deep neural framework to generate thumbnails of any size and aspect ratio,…