8 citations · 9 across the 3 of their papers we have counts for
8 papers
FaIRCoP: Facial Image Retrieval using Contrastive Personalization
Devansh Gupta, Aditya Saini, Drishti Bhasin +5
Retrieving facial images from attributes plays a vital role in various systems such as face recognition and suspect identification. Compared to other image retrieval tasks, facial…
Multimodal Research in Vision and Language: A Review of Current and Emerging Trends
Shagun Uppal, Sarthak Bhagat, Devamanyu Hazarika +4
Deep Learning and its applications have cascaded impactful research and development with a diverse range of modalities present in the real-world data. More recently, this has enhan…
DisCont: Self-Supervised Visual Attribute Disentanglement using Context Vectors
Sarthak Bhagat, Vishaal Udandarao, Shagun Uppal
Disentangling the underlying feature attributes within an image with no prior supervision is a challenging task. Models that can disentangle attributes well provide greater interpr…
C3VQG: Category Consistent Cyclic Visual Question Generation
Shagun Uppal, Anish Madan, Sarthak Bhagat +2
Visual Question Generation (VQG) is the task of generating natural questions based on an image. Popular methods in the past have explored image-to-sequence architectures trained wi…
Disentangling Multiple Features in Video Sequences using Gaussian Processes in Variational Autoencoders
Sarthak Bhagat, Shagun Uppal, Zhuyun Yin +1
We introduce MGP-VAE (Multi-disentangled-features Gaussian Processes Variational AutoEncoder), a variational autoencoder which uses Gaussian processes (GP) to model the latent spac…
Learning based Methods for Code Runtime Complexity Prediction
Jagriti Sikka, Kushal Satya, Yaman Kumar +3
Predicting the runtime complexity of a programming code is an arduous task. In fact, even for humans, it requires a subtle analysis and comprehensive knowledge of algorithms to pre…