1 citations · 5 across the 34 of their papers we have counts for
36 papers
RoboFind: Multi-Agent Personalized Object Search for People Who Are Blind or Have Low Vision
Ruiping Liu, Shaofang Quan, Qian Yin +10
Blind and low-vision users often need to locate a specific personal object rather than an arbitrary instance of the same category. The task calls for a robot that can move through…
XLocalizer: Cross-grained Alignment for Progressive Cross-view Video Geo-localization
Zichao Zeng, Weijia Fan, Yufan Chen +7
Cross-view Video Geo-localization (CVG) aims to localize ground-view videos by retrieving their corresponding geo-tagged aerial images. However, CVG approaches rely on fixed-length…
GuideFetch: A Task Coordination Framework for Concurrent Navigation and Object Retrieval in Assistive Robot Dogs
Qian Yin, Ruiping Liu, Kunyu Peng +7
Consider one robot guide dog escorting a blind user to a seat while a second retrieves and delivers an object. We introduce \textsc{GuideFetch}, a framework for coordinating this c…
SeqLoc: Beyond the Single Frame for Cross-View Geo-Localization in Feature-Sparse Scenes
Junwei Zheng, Yun Huang, Ruize Dai +8
Cross-View Geo-Localization (CVGL) with OpenStreetMap (OSM) performs well in structure-rich urban environments but collapses in feature-sparse scenes such as rural roads. To study…
Faster or Stronger: Towards Flexible Visual Place Recognition via Weighted Aggregation and Token Pruning
Zichao Zeng, June Moh Goo, Junwei Zheng +4
Visual Place Recognition (VPR) aims to match a query image to reference images of the same place in a large-scale database. Recent state-of-the-art methods employ Vision Transforme…
TriALS: Triphasic-Aided Liver Lesion Segmentation Benchmark in Non-Contrast CT
Marawan Elbatel, Mohamed Ghonim, Jiaji Mao +62
Automated segmentation of liver lesions on non-contrast computed tomography (NCCT) is clinically important but fundamentally challenging, particularly in low-resource settings acro…