88 citations · 192 across the 9 of their papers we have counts for
15 papers · 1 filter
Learnable Motion-Focused Tokenization for Effective and Efficient Video Unsupervised Domain Adaptation
Tzu Ling Liu, Ian Stavness, Mrigank Rochan
Video Unsupervised Domain Adaptation (VUDA) poses a significant challenge in action recognition, requiring the adaptation of a model from a labeled source domain to an unlabeled ta…
Test-Time Adaptation for Video Highlight Detection Using Meta-Auxiliary Learning and Cross-Modality Hallucinations
Zahidul Islam, Sujoy Paul, Mrigank Rochan
Existing video highlight detection methods, although advanced, struggle to generalize well to all test videos. These methods typically employ a generic highlight detection model fo…
Referring Segmentation in Images and Videos with Cross-Modal Self-Attention Network
Linwei Ye, Mrigank Rochan, Zhi Liu +2
We consider the problem of referring segmentation in images and videos with natural language. Given an input image (or video) and a referring expression, the goal is to segment the…
AdaCrowd: Unlabeled Scene Adaptation for Crowd Counting
Mahesh Kumar Krishna Reddy, Mrigank Rochan, Yiwei Lu +1
We address the problem of image-based crowd counting. In particular, we propose a new problem called unlabeled scene-adaptive crowd counting. Given a new target scene, we would lik…
Sentence Guided Temporal Modulation for Dynamic Video Thumbnail Generation
Mrigank Rochan, Mahesh Kumar Krishna Reddy, Yang Wang
We consider the problem of sentence specified dynamic video thumbnail generation. Given an input video and a user query sentence, the goal is to generate a video thumbnail that not…
Adaptive Video Highlight Detection by Learning from User History
Mrigank Rochan, Mahesh Kumar Krishna Reddy, Linwei Ye +1
Recently, there is an increasing interest in highlight detection research where the goal is to create a short duration video from a longer video by extracting its interesting momen…