3 papers
cs.RO2026
TemporalFlow-VLA: Learning Physically Grounded Execution History for Long-Horizon Robot Manipulation
Jiarui Yang, Yehao Lu, Yuning Su +9
Vision-language-action (VLA) models leverage pretrained vision-language representations for robot control, yet simply adding historical frames does not reliably capture recent phys…
cs.CV2026
V-Link: Recovering Lost Visual Representations in Action DiT for Vision-Language-Action Models
Yehao Lu, Jiarui Yang, Yuning Su +10
Vision-language-action (VLA) models provide a scalable path toward generalist robotic manipulation by integrating visual perception, language understanding, and continuous action c…
cs.LG2026
PolyFlow: Safe and Efficient Polytope-Constrained Flow Matching with Constraint Embedding and Projection-free Update
Jianming Ma, Qiyue Yang, Yang Zhang +4
While flow-based generative models have demonstrated strong performance across a wide range of domains, deploying them in safety-critical physical systems remains challenging due t…