2 papers
cs.RO2026
Habilis-: A Fast-Motion and Long-Lasting On-Device Vision-Language-Action Model
Tommoro Robotics, :, Jesoon Kang +22
We introduce Habilis-, a fast-motion and long-lasting on-device vision-language-action (VLA) model designed for real-world deployment. Current VLA evaluation remains largely co…
cs.CL2025
Towards LLM-Centric Multimodal Fusion: A Survey on Integration Strategies and Techniques
Jisu An, Junseok Lee, Jeoungeun Lee +1
The rapid progress of Multimodal Large Language Models(MLLMs) has transformed the AI landscape. These models combine pre-trained LLMs with various modality encoders. This integrati…