Showing 2025Show all
2 papers · 1 filter
cs.CV2025
Zero-Shot 3D Visual Grounding from Vision-Language Models
Rong Li, Shijie Li, Lingdong Kong +2
3D Visual Grounding (3DVG) seeks to locate target objects in 3D scenes using natural language descriptions, enabling downstream applications such as augmented reality and robotics.…
cs.RO2025
Stairway to Success: An Online Floor-Aware Zero-Shot Object-Goal Navigation Framework via LLM-Driven Coarse-to-Fine Exploration
Zeying Gong, Rong Li, Tianshuai Hu +6
Deployable service and delivery robots struggle to navigate multi-floor buildings to reach object goals, as existing systems fail due to single-floor assumptions and requirements f…