4 papers
VISOR: VIsual Spatial Object Reasoning for Language-driven Object Navigation
Francesco Taioli, Shiping Yang, Sonia Raychaudhuri +3
Language-driven object navigation requires agents to interpret natural language descriptions of target objects, which combine intrinsic and extrinsic attributes for instance recogn…
MLFM: Multi-Layered Feature Maps for Richer Language Understanding in Zero-Shot Semantic Navigation
Sonia Raychaudhuri, Enrico Cancelli, Tommaso Campari +3
Recent progress in large vision-language models has driven improvements in language-based semantic navigation, where an embodied agent must reach a target object described in natur…
Semantic Mapping in Indoor Embodied AI -- A Survey on Advances, Challenges, and Future Directions
Sonia Raychaudhuri, Angel X. Chang
Intelligent embodied agents (e.g. robots) need to perform complex semantic tasks in unfamiliar environments. Among many skills that the agents need to possess, building and maintai…
Zero-shot Object-Centric Instruction Following: Integrating Foundation Models with Traditional Navigation
Sonia Raychaudhuri, Duy Ta, Katrina Ashton +3
Large scale scenes such as multifloor homes can be robustly and efficiently mapped with a 3D graph of landmarks estimated jointly with robot poses in a factor graph, a technique co…