From the 1 of 5 linked papers with an AI index.
5 papers
Towards Human-like Physical Intelligence: Lifelong Vision-Language-Action Learning for Robotic Manipulation
Yao He, Gan Sun, Wenqi Liang +2
The paper introduces LifelongVLA, a framework that enables robots to continuously learn new manipulation tasks by using a dual-timescale adaptation mechanism and a cache-efficient…
Continual Hand-Eye Calibration for Open-world Robotic Manipulation
Fazeng Li, Gan Sun, Chenxi Liu +3
Hand-eye calibration through visual localization is a critical capability for robotic manipulation in open-world environments. However, most deep learning-based calibration models…
Federated Multi-Task Clustering
Suyan Dai, Gan Sun, Fazeng Li +3
Spectral clustering has emerged as one of the most effective clustering algorithms due to its superior performance. However, most existing models are designed for centralized setti…
PixelVLA: Advancing Pixel-level Understanding in Vision-Language-Action Model
Wenqi Liang, Gan Sun, Yao He +5
Vision-Language-Action models (VLAs) are emerging as powerful tools for learning generalizable visuomotor control policies. However, current VLAs are mostly trained on large-scale…
OpenVLN: Open-world Aerial Vision-Language Navigation
Peican Lin, Gan Sun, Chenxi Liu +3
Vision-language models (VLMs) have been widely-applied in ground-based vision-language navigation (VLN). However, the vast complexity of outdoor aerial environments compounds data…