33 citations · 37 across the 19 of their papers we have counts for
3 papers · 1 filter
Gravity-Aware Monocular 3D Human-Object Reconstruction
Rishabh Dabral, Soshi Shimada, Arjun Jain +2
This paper proposes GraviCap, i.e., a new approach for joint markerless 3D human motion capture and object trajectory estimation from monocular RGB videos. We focus on scenes with…
Cross-Modal learning for Audio-Visual Video Parsing
Jatin Lamba, Abhishek, Jayaprakash Akula +3
In this paper, we present a novel approach to the audio-visual video parsing (AVVP) task that demarcates events from a video separately for audio and visual modalities. The propose…
Rudder: A Cross Lingual Video and Text Retrieval Dataset
Jayaprakash A, Abhishek, Rishabh Dabral +2
Video retrieval using natural language queries requires learning semantically meaningful joint embeddings between the text and the audio-visual input. Often, such joint embeddings…