Showing cs.CVShow all
2 papers · 1 filter
cs.CV2023
VQA with Cascade of Self- and Co-Attention Blocks
Aakansha Mishra, Ashish Anand, Prithwijit Guha
The use of complex attention modules has improved the performance of the Visual Question Answering (VQA) task. This work aims to learn an improved multi-modal representation throug…
cs.CV2015
An Occlusion Reasoning Scheme for Monocular Pedestrian Tracking in Dynamic Scenes
Sourav Garg, Swagat Kumar, Rajesh Ratnakaram +1
This paper looks into the problem of pedestrian tracking using a monocular, potentially moving, uncalibrated camera. The pedestrians are located in each frame using a standard huma…