2 papers
cs.RO2026
Retrieve in Time, Correct in Frequency
Yuze Fan, Yue Cao, Pengjie Gao +7
Frozen vision-language-action (VLA) policies generate temporally extended action chunks, but long-horizon manipulation remains vulnerable to accumulated execution error and visual…
cs.RO2025
DiffE2E: Rethinking End-to-End Driving with a Hybrid Action Diffusion and Supervised Policy
Rui Zhao, Yuze Fan, Ziguo Chen +2
End-to-end learning has emerged as a transformative paradigm in autonomous driving. However, the inherently multimodal nature of driving behaviors and the generalization challenges…