5 papers
Learnable Query Aggregation with KV Routing for Cross-view Geo-localisation
Hualin Ye, Bingxi Liu, Jixiang Du +3
Cross-view geo-localisation (CVGL) aims to estimate the geographic location of a query image by matching it with images from a large-scale database. However, the significant view-p…
TrackingMiM: Efficient Mamba-in-Mamba Serialization for Real-time UAV Object Tracking
Bingxi Liu, Calvin Chen, Junhao Li +5
The Vision Transformer (ViT) model has long struggled with the challenge of quadratic complexity, a limitation that becomes especially critical in unmanned aerial vehicle (UAV) tra…
EmbodiedPlace: Learning Mixture-of-Features with Embodied Constraints for Visual Place Recognition
Bingxi Liu, Hao Chen, Shiyi Guo +3
Visual Place Recognition (VPR) is a scene-oriented image retrieval problem in computer vision in which re-ranking based on local features is commonly employed to improve performanc…
SuperPlace: The Renaissance of Classical Feature Aggregation for Visual Place Recognition in the Era of Foundation Models
Bingxi Liu, Pengju Zhang, Li He +5
Recent visual place recognition (VPR) approaches have leveraged foundation models (FM) and introduced novel aggregation techniques. However, these methods have failed to fully expl…
TextInPlace: Indoor Visual Place Recognition in Repetitive Structures with Scene Text Spotting and Verification
Huaqi Tao, Bingxi Liu, Calvin Chen +4
Visual Place Recognition (VPR) is a crucial capability for long-term autonomous robots, enabling them to identify previously visited locations using visual information. However, ex…