3 papers
cs.RO2025
Bridging Scale Discrepancies in Robotic Control via Language-Based Action Representations
Yuchi Zhang, Churui Sun, Shiqi Liang +4
Recent end-to-end robotic manipulation research increasingly adopts architectures inspired by large language models to enable robust manipulation. However, a critical challenge ari…
cs.RO2025
iFlyBot-VLM Technical Report
Xin Nie, Zhiyuan Cheng, Yuan Zhang +4
We introduce iFlyBot-VLM, a general-purpose Vision-Language Model (VLM) used to improve the domain of Embodied Intelligence. The central objective of iFlyBot-VLM is to bridge the c…
cs.CV2025
iFlyBot-VLA Technical Report
Yuan Zhang, Chenyu Xue, Wenjie Xu +3
We introduce iFlyBot-VLA, a large-scale Vision-Language-Action (VLA) model trained under a novel framework. The main contributions are listed as follows: (1) a latent action model…