activity
20162025
most citedSipMask: Spatial Information Preservation for Fast Image and Video Instance Segmentation

28 citations · 66 across the 12 of their papers we have counts for

collaborators

15 papers

cs.CV20221 cited

PS-ARM: An End-to-End Attention-aware Relation Mixer Network for Person Search

Mustansar Fiaz, Hisham Cholakkal, Sanath Narayan +2

Person search is a challenging problem with various real-world applications, that aims at joint person detection and re-identification of a query person from uncropped gallery imag…

cs.CV2022

CMR3D: Contextualized Multi-Stage Refinement for 3D Object Detection

Dhanalaxmi Gaddam, Jean Lahoud, Fahad Shahbaz Khan +2

Existing deep learning-based 3D object detectors typically rely on the appearance of individual objects and do not explicitly pay attention to the rich contextual information of th…

cs.CV2022

PSTR: End-to-End One-Step Person Search With Transformers

Jiale Cao, Yanwei Pang, Rao Muhammad Anwer +4

We propose a novel one-step transformer-based person search framework, PSTR, that jointly performs person detection and re-identification (re-id) in a single architecture. PSTR com…

cs.CV20222 cited

Video Instance Segmentation via Multi-scale Spatio-temporal Split Attention Transformer

Omkar Thawakar, Sanath Narayan, Jiale Cao +6

State-of-the-art transformer-based video instance segmentation (VIS) approaches typically utilize either single-scale spatio-temporal features or per-frame multi-scale features dur…

cs.CV20212 cited

Structured Latent Embeddings for Recognizing Unseen Classes in Unseen Domains

Shivam Chandhok, Sanath Narayan, Hisham Cholakkal +4

The need to address the scarcity of task-specific annotated data has resulted in concerted efforts in recent years for specific settings such as zero-shot learning (ZSL) and domain…

cs.CV2021

Handwriting Transformers

Ankan Kumar Bhunia, Salman Khan, Hisham Cholakkal +3

We propose a novel transformer-based styled handwritten text image generation approach, HWT, that strives to learn both style-content entanglement as well as global and local writi…