12 citations · 12 across the 3 of their papers we have counts for
Showing cs.ROShow all
2 papers · 1 filter
cs.RO2026
How Should Vision-Language-Action Models Use Proprioceptive State?
Yiren Zhao, Ziyang Chen, Ziyang Rao +5
Recent Vision-Language-Action (VLA) models almost universally take robot proprioceptive state as input, yet wire it in incompatible ways -- serialized into text prompts, projected…
cs.RO2026
Source-Lifted Flow Matching for Intervenable Multimodal Imitation
He Zhang, Ying Sun, Pengteng Li +6
Flow-matching policies are promising for imitation learning because they model complex multimodal action distributions. However, their stochasticity is largely passive: repeated sa…