activity
20182022
most citedIn Defense of Pseudo-Labeling: An Uncertainty-Aware Pseudo-label Selection Framework for Semi-Supervised Learning

73 citations · 82 across the 10 of their papers we have counts for

collaborators

19 papers

cs.CV2022

Video Action Detection: Analysing Limitations and Challenges

Rajat Modi, Aayush Jung Rana, Akash Kumar +4

Beyond possessing large enough size to feed data hungry machines (eg, transformers), what attributes measure the quality of a dataset? Assuming that the definitions of such attribu…

cs.CV2021

LARNet: Latent Action Representation for Human Action Synthesis

Naman Biyani, Aayush J Rana, Shruti Vyas +1

We present LARNet, a novel end-to-end approach for generating human action videos. A joint generative modeling of appearance and dynamics to synthesize a video is very challenging…

cs.CV20211 cited

"Knights": First Place Submission for VIPriors21 Action Recognition Challenge at ICCV 2021

Ishan Dave, Naman Biyani, Brandon Clark +3

This technical report presents our approach "Knights" to solve the action recognition task on a small subset of Kinetics-400 i.e. Kinetics400ViPriors without using any extra-data.…

cs.MM2021

NoisyActions2M: A Multimedia Dataset for Video Understanding from Noisy Labels

Mohit Sharma, Raj Patra, Harshal Desai +3

Deep learning has shown remarkable progress in a wide range of problems. However, efficient training of such models requires large-scale datasets, and getting annotations for such…

cs.CV20216 cited

TinyAction Challenge: Recognizing Real-world Low-resolution Activities in Videos

Praveen Tirupattur, Aayush J Rana, Tushar Sangam +3

This paper summarizes the TinyAction challenge which was organized in ActivityNet workshop at CVPR 2021. This challenge focuses on recognizing real-world low-resolution activities…

cs.CV2021

Novel View Video Prediction Using a Dual Representation

Sarah Shiraz, Krishna Regmi, Shruti Vyas +2

We address the problem of novel view video prediction; given a set of input video clips from a single/multiple views, our network is able to predict the video from a novel view. Th…