4 citations · 8 across the 2 of their papers we have counts for
4 papers
A Generic Visualization Approach for Convolutional Neural Networks
Ahmed Taha, Xitong Yang, Abhinav Shrivastava +1
Retrieval networks are essential for searching and indexing. Compared to classification networks, attention visualization for retrieval networks is hardly studied. We formulate att…
STEP: Spatio-Temporal Progressive Learning for Video Action Detection
Xitong Yang, Xiaodong Yang, Ming-Yu Liu +3
In this paper, we propose Spatio-TEmporal Progressive (STEP) action detector---a progressive learning framework for spatio-temporal action detection in videos. Starting from a hand…
Exploring Uncertainty in Conditional Multi-Modal Retrieval Systems
Ahmed Taha, Yi-Ting Chen, Xitong Yang +2
We cast visual retrieval as a regression problem by posing triplet loss as a regression loss. This enables epistemic uncertainty estimation using dropout as a Bayesian approximatio…
Two Stream Self-Supervised Learning for Action Recognition
Ahmed Taha, Moustafa Meshry, Xitong Yang +2
We present a self-supervised approach using spatio-temporal signals between video frames for action recognition. A two-stream architecture is leveraged to tangle spatial and tempor…