51 citations · 104 across the 5 of their papers we have counts for
6 papers
Sentence Guided Temporal Modulation for Dynamic Video Thumbnail Generation
Mrigank Rochan, Mahesh Kumar Krishna Reddy, Yang Wang
We consider the problem of sentence specified dynamic video thumbnail generation. Given an input video and a user query sentence, the goal is to generate a video thumbnail that not…
Adaptive Video Highlight Detection by Learning from User History
Mrigank Rochan, Mahesh Kumar Krishna Reddy, Linwei Ye +1
Recently, there is an increasing interest in highlight detection research where the goal is to create a short duration video from a longer video by extracting its interesting momen…
Few-Shot Scene Adaptive Crowd Counting Using Meta-Learning
Mahesh Kumar Krishna Reddy, Mohammad Hossain, Mrigank Rochan +1
We consider the problem of few-shot scene adaptive crowd counting. Given a target camera scene, our goal is to adapt a model to this specific scene with only a few labeled images o…
Convolutional Temporal Attention Model for Video-based Person Re-identification
Tanzila Rahman, Mrigank Rochan, Yang Wang
The goal of video-based person re-identification is to match two input videos, so that the distance of the two videos is small if two videos contain the same person. A common appro…
Cross-Modal Self-Attention Network for Referring Image Segmentation
Linwei Ye, Mrigank Rochan, Zhi Liu +1
We consider the problem of referring image segmentation. Given an input image and a natural language expression, the goal is to segment the object referred by the language expressi…
Label Refinement Network for Coarse-to-Fine Semantic Segmentation
Md Amirul Islam, Shujon Naha, Mrigank Rochan +2
We consider the problem of semantic image segmentation using deep convolutional neural networks. We propose a novel network architecture called the label refinement network that pr…