works on

From the 1 of 5 linked papers with an AI index.

collaborators

5 papers

cs.CV2026

Breaking Déjà Vu: Independent Auditing of Visual Place Recognition through Vision-Language Reasoning

Sania Waheed, Michael Milford, Sarvapali D. Ramchurn +1

The paper proposes an independent auditing framework that uses vision‑language models to verify visual place‑recognition matches, improving recall and reducing false acceptances wi…

cs.CV2026

VLM-Guided Visual Place Recognition for Planet-Scale Geo-Localization

Sania Waheed, Na Min An, Michael Milford +2

Geo-localization from a single image at planet scale (essentially an advanced or extreme version of the kidnapped robot problem) is a fundamental and challenging task in applicatio…

cs.CV2026

Image-based Geo-localization for Robotics: Are Black-box Vision-Language Models there yet?

Sania Waheed, Bruno Ferrarini, Michael Milford +2

The advances in Vision-Language models (VLMs) offer exciting opportunities for robotic applications involving image geo-localization - the problem of identifying the geo-coordinate…

cs.CV2025

Image Embedding Sampling Method for Diverse Captioning

Sania Waheed, Na Min An

Image Captioning for state-of-the-art VLMs has significantly improved over time; however, this comes at the cost of increased computational complexity, making them less accessible…

cs.CV2025

Assessing the Geolocation Capabilities, Limitations and Societal Risks of Generative Vision-Language Models

Oliver Grainge, Sania Waheed, Jack Stilgoe +2

Geo-localization is the task of identifying the location of an image using visual cues alone. It has beneficial applications, such as improving disaster response, enhancing navigat…