5 papers
Foresight: Iterative Reasoning About Clues that Matter for Navigation
Arthur Zhang, Carl Qi, Donne Su +3
Open-world mapless navigation from sparse language instructions requires resolving underspecified goals and inferring which environmental cues are relevant for reaching the goal. F…
GuideTWSI: A Diverse Tactile Walking Surface Indicator Dataset from Synthetic and Real-World Images for Blind and Low-Vision Navigation
Hochul Hwang, Soowan Yang, Anh N. H. Nguyen +6
Tactile Walking Surface Indicators (TWSIs) are safety-critical landmarks that blind and low-vision (BLV) pedestrians use to locate crossings and hazard zones. From our observation…
VENTURA: Adapting Image Diffusion Models for Unified Task Conditioned Navigation
Arthur Zhang, Xiangyun Meng, Luca Calliari +5
Robots must adapt to diverse human instructions and operate safely in unstructured, open-world environments. Recent Vision-Language models (VLMs) offer strong priors for grounding…
CREStE: Scalable Mapless Navigation with Internet Scale Priors and Counterfactual Guidance
Arthur Zhang, Harshit Sikchi, Amy Zhang +1
We introduce CREStE, a scalable learning-based mapless navigation framework to address the open-world generalization and robustness challenges of outdoor urban navigation. Key to a…
Guiding Diffusion with Deep Geometric Moments: Balancing Fidelity and Variation
Sangmin Jung, Utkarsh Nath, Yezhou Yang +5
Text-to-image generation models have achieved remarkable capabilities in synthesizing images, but often struggle to provide fine-grained control over the output. Existing guidance…