most citedBridging Zero-shot Object Navigation and Foundation Models through Pixel-Guided Navigation Skill

4 citations · 12 across the 5 of their papers we have counts for

collaborators

5 papers

cs.RO20243 cited

InstructNav: Zero-shot System for Generic Instruction Navigation in Unexplored Environment

Yuxing Long, Wenzhe Cai, Hongcheng Wang +2

Enabling robots to navigate following diverse language instructions in unexplored environments is an attractive goal for human-robot interaction. However, this goal is challenging…

cs.AI20244 cited

Empowering Large Language Models on Robotic Manipulation with Affordance Prompting

Guangran Cheng, Chuheng Zhang, Wenzhe Cai +3

While large language models (LLMs) are successful in completing various language processing tasks, they easily fail to interact with the physical world by generating control sequen…

cs.RO2023

Robust Navigation with Cross-Modal Fusion and Knowledge Transfer

Wenzhe Cai, Guangran Cheng, Lingyue Kong +2

Recently, learning-based approaches show promising results in navigation tasks. However, the poor generalization capability and the simulation-reality gap prevent a wide range of a…

cs.RO20234 cited

Bridging Zero-shot Object Navigation and Foundation Models through Pixel-Guided Navigation Skill

Wenzhe Cai, Siyuan Huang, Guangran Cheng +4

Zero-shot object navigation is a challenging task for home-assistance robots. This task emphasizes visual grounding, commonsense inference and locomotion abilities, where the first…

cs.RO20231 cited

Discuss Before Moving: Visual Language Navigation via Multi-expert Discussions

Yuxing Long, Xiaoqi Li, Wenzhe Cai +1

Visual language navigation (VLN) is an embodied task demanding a wide range of skills encompassing understanding, perception, and planning. For such a multifaceted challenge, previ…