12 citations · 12 across the 3 of their papers we have counts for
3 papers
cs.RO2026
How Should Vision-Language-Action Models Use Proprioceptive State?
Yiren Zhao, Ziyang Chen, Ziyang Rao +5
Recent Vision-Language-Action (VLA) models almost universally take robot proprioceptive state as input, yet wire it in incompatible ways -- serialized into text prompts, projected…
cs.RO2026
Source-Lifted Flow Matching for Intervenable Multimodal Imitation
He Zhang, Ying Sun, Pengteng Li +6
Flow-matching policies are promising for imitation learning because they model complex multimodal action distributions. However, their stochasticity is largely passive: repeated sa…
cs.MM2023★ 12 cited
Interactive Interior Design Recommendation via Coarse-to-fine Multimodal Reinforcement Learning
He Zhang, Ying Sun, Weiyu Guo +4
Personalized interior decoration design often incurs high labor costs. Recent efforts in developing intelligent interior design systems have focused on generating textual requireme…