1 citations · 2 across the 2 of their papers we have counts for
Showing cs.CVShow all
3 papers · 1 filter
cs.CV2023★ 1 cited
Understanding Video Transformers for Segmentation: A Survey of Application and Interpretability
Rezaul Karim, Richard P. Wildes
Video segmentation encompasses a wide range of categories of problem formulation, e.g., object, scene, actor-action and multimodal video segmentation, for delineating task-specific…
cs.CV2023★ 1 cited
Multiscale Memory Comparator Transformer for Few-Shot Video Segmentation
Mennatullah Siam, Rezaul Karim, He Zhao +1
Few-shot video segmentation is the task of delineating a specific novel class in a query video using few labelled support images. Typical approaches compare support and query featu…
cs.CV2023
MED-VT++: Unifying Multimodal Learning with a Multiscale Encoder-Decoder Video Transformer
Rezaul Karim, He Zhao, Richard P. Wildes +1
In this paper, we present an end-to-end trainable unified multiscale encoder-decoder transformer that is focused on dense prediction tasks in video. The presented Multiscale Encode…