activity
20162026
most citedSigNet: Convolutional Siamese Network for Writer Independent Offline Signature Verification

179 citations · 188 across the 6 of their papers we have counts for

collaborators
Showing cs.CVShow all

11 papers · 1 filter

cs.CV2026

DocRevive: A Unified Pipeline for Document Text Restoration

Kunal Purkayastha, Ayan Banerjee, Josep Llados +1

In Document Understanding, the challenge of reconstructing damaged, occluded, or incomplete text remains a critical yet unexplored problem. Subsequent document understanding tasks…

cs.CV2025

NoTeS-Bank: Benchmarking Neural Transcription and Search for Scientific Notes Understanding

Aniket Pal, Sanket Biswas, Alloy Das +6

Understanding and reasoning over academic handwritten notes remains a challenge in document AI, particularly for mathematical equations, diagrams, and scientific notations. Existin…

cs.CV20213 cited

Localizing Infinity-shaped fishes: Sketch-guided object localization in the wild

Pau Riba, Sounak Dey, Ali Furkan Biten +1

This work investigates the problem of sketch-guided object localization (SGOL), where human sketches are used as queries to conduct the object localization in natural images. In th…

cs.CV2019

Doodle to Search: Practical Zero-Shot Sketch-based Image Retrieval

Sounak Dey, Pau Riba, Anjan Dutta +2

In this paper, we investigate the problem of zero-shot sketch-based image retrieval (ZS-SBIR), where human sketches are used as queries to conduct retrieval of photos from unseen c…

cs.CV2018

Hierarchical stochastic graphlet embedding for graph-based pattern recognition

Anjan Dutta, Pau Riba, Josep Lladós +1

Despite being very successful within the pattern recognition and machine learning community, graph-based methods are often unusable because of the lack of mathematical operations d…

cs.CV2018

Learning Cross-Modal Deep Embeddings for Multi-Object Image Retrieval using Text and Sketch

Sounak Dey, Anjan Dutta, Suman K. Ghosh +3

In this work we introduce a cross modal image retrieval system that allows both text and sketch as input modalities for the query. A cross-modal deep network architecture is formul…