1 paper · 1 filter
Christian Rauch, Björn Ellensohn, Linus Nwankwo +2
Semantic scene understanding in robotics requires representations that are both metric-accurate and queryable via natural language in real-time. While recent Vision-Language Models…