17 citations · 27 across the 5 of their papers we have counts for
5 papers
Towards Zero-shot Cross-lingual Image Retrieval and Tagging
Pranav Aggarwal, Ritiz Tambi, Ajinkya Kale
There has been a recent spike in interest in multi-modal Language and Vision problems. On the language side, most of these models primarily focus on English since most multi-modal…
Towards Zero-shot Cross-lingual Image Retrieval
Pranav Aggarwal, Ajinkya Kale
There has been a recent spike in interest in multi-modal Language and Vision problems. On the language side, most of these models primarily focus on English since most multi-modal…
Multi-Modal Retrieval using Graph Neural Networks
Aashish Kumar Misraa, Ajinkya Kale, Pranav Aggarwal +1
Most real world applications of image retrieval such as Adobe Stock, which is a marketplace for stock photography and illustrations, need a way for users to find images which are b…
Multitask Text-to-Visual Embedding with Titles and Clickthrough Data
Pranav Aggarwal, Zhe Lin, Baldo Faieta +1
Text-visual (or called semantic-visual) embedding is a central problem in vision-language research. It typically involves mapping of an image and a text description to a common fea…
A Deep Learning Approach to Drone Monitoring
Yueru Chen, Pranav Aggarwal, Jongmoo Choi +1
A drone monitoring system that integrates deep-learning-based detection and tracking modules is proposed in this work. The biggest challenge in adopting deep learning methods for d…