FASTNav: Fine-tuned Adaptive Small-language-models Trained for Multi-point Robot Navigation
arXiv:2411.13262 · doi:10.1109/LRA.2024.3506280
Abstract
With the rapid development of large language models (LLM), robots are starting to enjoy the benefits of new interaction methods that large language models bring. Because edge computing fulfills the needs for rapid response, privacy, and network autonomy, we believe it facilitates the extensive deployment of large models for robot navigation across various industries. To enable local deployment of language models on edge devices, we adopt some model boosting methods. In this paper, we propose FASTNav - a method for boosting lightweight LLMs, also known as small language models (SLMs), for robot navigation. The proposed method contains three modules: fine-tuning, teacher-student iteration, and language-based multi-point robot navigation. We train and evaluate models with FASTNav in both simulation and real robots, proving that we can deploy them with low cost, high accuracy and low response time. Compared to other model compression methods, FASTNav shows potential in the local deployment of language models and tends to be a promising solution for language-guided robot navigation on edge devices.
References in corpus (9)
- Do As I Can, Not As I Say: Grounding Language in Robotic Affordances
- The Marathon 2: A Navigation System
- LLM+P: Empowering Large Language Models with Optimal Planning Proficiency
- LLM-Pruner: On the Structural Pruning of Large Language Models
- Orca 2: Teaching Small Language Models How to Reason
- In-context Learning Distillation: Transferring Few-shot Learning Ability of Pre-trained Language Models
- ZeroQuant-FP: A Leap Forward in LLMs Post-Training W4A8 Quantization Using Floating-Point Formats
- Unifying Large Language Model and Deep Reinforcement Learning for Human-in-Loop Interactive Socially-aware Navigation
- Towards the Law of Capacity Gap in Distilling Language Models