From the 1 of 5 linked papers with an AI index.
5 papers
Breaking Déjà Vu: Independent Auditing of Visual Place Recognition through Vision-Language Reasoning
Sania Waheed, Michael Milford, Sarvapali D. Ramchurn +1
The paper proposes an independent auditing framework that uses vision‑language models to verify visual place‑recognition matches, improving recall and reducing false acceptances wi…
VLM-Guided Visual Place Recognition for Planet-Scale Geo-Localization
Sania Waheed, Na Min An, Michael Milford +2
Geo-localization from a single image at planet scale (essentially an advanced or extreme version of the kidnapped robot problem) is a fundamental and challenging task in applicatio…
Image-based Geo-localization for Robotics: Are Black-box Vision-Language Models there yet?
Sania Waheed, Bruno Ferrarini, Michael Milford +2
The advances in Vision-Language models (VLMs) offer exciting opportunities for robotic applications involving image geo-localization - the problem of identifying the geo-coordinate…
Image Embedding Sampling Method for Diverse Captioning
Sania Waheed, Na Min An
Image Captioning for state-of-the-art VLMs has significantly improved over time; however, this comes at the cost of increased computational complexity, making them less accessible…
Assessing the Geolocation Capabilities, Limitations and Societal Risks of Generative Vision-Language Models
Oliver Grainge, Sania Waheed, Jack Stilgoe +2
Geo-localization is the task of identifying the location of an image using visual cues alone. It has beneficial applications, such as improving disaster response, enhancing navigat…