41 citations · 41 across the 2 of their papers we have counts for
2 papers
cs.SD2024
Learning Spatially-Aware Language and Audio Embeddings
Bhavika Devnani, Skyler Seto, Zakaria Aldeneh +5
Humans can picture a sound scene given an imprecise natural language description. For example, it is easy to imagine an acoustic environment given a phrase like "the lion roar came…
cs.CV2022★ 41 cited
ZSON: Zero-Shot Object-Goal Navigation using Multimodal Goal Embeddings
Arjun Majumdar, Gunjan Aggarwal, Bhavika Devnani +2
We present a scalable approach for learning open-world object-goal navigation (ObjectNav) -- the task of asking a virtual robot (agent) to find any instance of an object in an unex…