Showing cs.ROShow all
2 papers · 1 filter
cs.RO2026
TFP: Temporally Conditioned Memory-Fusion Policies for Visuomotor Learning
Yushen Liang, Yue Peng, Baosheng Jin +6
Vision--Language--Action (VLA) policies such as and OpenVLA perform well on many manipulation tasks, but they are often reactive: the next action is predicted from the c…
cs.RO2026
GVLA: Geometric inductive bias for Vision-Language-Action Models
Yue Peng, Yongzhe Zhao, Artur Habuda +5
Vision-language-action (VLA) models have made rapid progress in generalist robot manipulation by harnessing semantic knowledge from pretrained vision-language backbones, but their…