65 citations · 130 across the 14 of their papers we have counts for
4 papers · 1 filter
Nose, eyes and ears: Head pose estimation by locating facial keypoints
Aryaman Gupta, Kalpit Thakkar, Vineet Gandhi +1
Monocular head pose estimation requires learning a model that computes the intrinsic Euler angles for pose (yaw, pitch, roll) from an input image of human face. Annotating ground t…
Watch to Edit: Video Retargeting using Gaze
Kranthi Kumar, Moneish Kumar, Vineet Gandhi +1
We present a novel approach to optimally retarget videos for varied displays with differing aspect ratios by preserving salient scene content discovered via eye tracking. Our algor…
MergeNet: A Deep Net Architecture for Small Obstacle Discovery
Krishnam Gupta, Syed Ashar Javed, Vineet Gandhi +1
We present here, a novel network architecture called MergeNet for discovering small obstacles for on-road scenes in the context of autonomous driving. The basis of the architecture…
Learning Unsupervised Visual Grounding Through Semantic Self-Supervision
Syed Ashar Javed, Shreyas Saxena, Vineet Gandhi
Localizing natural language phrases in images is a challenging problem that requires joint understanding of both the textual and visual modalities. In the unsupervised setting, lac…