8 citations · 9 across the 5 of their papers we have counts for
10 papers
Symbolic Graph Inference for Compound Scene Understanding
FNU Aryan, Simon Stepputtis, Sarthak Bhagat +4
Scene understanding is a fundamental capability needed in many domains, ranging from question-answering to robotics. Unlike recent end-to-end approaches that must explicitly learn…
FaIRCoP: Facial Image Retrieval using Contrastive Personalization
Devansh Gupta, Aditya Saini, Drishti Bhasin +5
Retrieving facial images from attributes plays a vital role in various systems such as face recognition and suspect identification. Compared to other image retrieval tasks, facial…
Target-Following Double Deep Q-Networks for UAVs
Sarthak Bhagat, P. B. Sujit
Target tracking in unknown real-world environments in the presence of obstacles and target motion uncertainty demand agents to develop an intrinsic understanding of the environment…
Multimodal Research in Vision and Language: A Review of Current and Emerging Trends
Shagun Uppal, Sarthak Bhagat, Devamanyu Hazarika +4
Deep Learning and its applications have cascaded impactful research and development with a diverse range of modalities present in the real-world data. More recently, this has enhan…
UAV Target Tracking in Urban Environments Using Deep Reinforcement Learning
Sarthak Bhagat, Sujit PB
Persistent target tracking in urban environments using UAV is a difficult task due to the limited field of view, visibility obstruction from obstacles and uncertain target motion.…
DisCont: Self-Supervised Visual Attribute Disentanglement using Context Vectors
Sarthak Bhagat, Vishaal Udandarao, Shagun Uppal
Disentangling the underlying feature attributes within an image with no prior supervision is a challenging task. Models that can disentangle attributes well provide greater interpr…