33 citations · 34 across the 2 of their papers we have counts for
4 papers
Rudder: A Cross Lingual Video and Text Retrieval Dataset
Jayaprakash A, Abhishek, Rishabh Dabral +2
Video retrieval using natural language queries requires learning semantically meaningful joint embeddings between the text and the audio-visual input. Often, such joint embeddings…
LIGHTEN: Learning Interactions with Graph and Hierarchical TEmporal Networks for HOI in videos
Sai Praneeth Reddy Sunkesula, Rishabh Dabral, Ganesh Ramakrishnan
Analyzing the interactions between humans and objects from a video includes identification of the relationships between humans and the objects present in the video. It can be thoug…
Multi-Person 3D Human Pose Estimation from Monocular Images
Rishabh Dabral, Nitesh B Gundavarapu, Rahul Mitra +3
Multi-person 3D human pose estimation from a single image is a challenging problem, especially for in-the-wild settings due to the lack of 3D annotated data. We propose HG-RCNN, a…
Progression Modelling for Online and Early Gesture Detection
Vikram Gupta, Sai Kumar Dwivedi, Rishabh Dabral +1
Online and Early detection of gestures is crucial for building touchless gesture based interfaces. These interfaces should operate on a stream of video frames instead of the comple…