37 citations · 174 across the 29 of their papers we have counts for
4 papers · 2 filters
Egocentric Video Task Translation
Zihui Xue, Yale Song, Kristen Grauman +1
Different video understanding tasks are typically treated in isolation, and even with distinct types of curated data (e.g., classifying sports in one dataset, tracking animals in a…
HistoPerm: A Permutation-Based View Generation Approach for Improving Histopathologic Feature Representation Learning
Joseph DiPalma, Lorenzo Torresani, Saeed Hassanpour
Deep learning has been effective for histology image analysis in digital pathology. However, many current deep learning approaches require large, strongly- or weakly-labeled images…
Deformable Video Transformer
Jue Wang, Lorenzo Torresani
Video transformers have recently emerged as an effective alternative to convolutional networks for action classification. However, most prior video transformers adopt either global…
Learning To Recognize Procedural Activities with Distant Supervision
Xudong Lin, Fabio Petroni, Gedas Bertasius +3
In this paper we consider the problem of classifying fine-grained, multi-step activities (e.g., cooking different recipes, making disparate home improvements, creating various form…