Showing cs.ROShow all
3 papers · 1 filter
cs.RO2026
PointACT: Vision-Language-Action Models with Multi-Scale Point-Action Interaction
Shizhe Chen, Paul Pacaud, Cordelia Schmid
Vision-Language-Action (VLA) models have shown strong potential for general-purpose robotic manipulation by leveraging large pretrained vision-language backbones. However, most exi…
cs.RO2026
MetricNet: Recovering Metric Scale in Generative Navigation Policies
Abhijeet Nayak, Débora Oliveira Makowski, Samiran Gode +2
Generative navigation policies have made rapid progress in improving end-to-end learned navigation. Despite their promising results, this paradigm has two structural problems. Firs…
cs.RO2025
FlowNav: Combining Flow Matching and Depth Priors for Efficient Navigation
Samiran Gode, Abhijeet Nayak, Débora N. P. Oliveira +3
Effective robot navigation in unseen environments is a challenging task that requires precise control actions at high frequencies. Recent advances have framed it as an image-goal-c…