3 papers
cs.CV2026
Traj-VLN: Learning Pixel-Space Interaction via Autoregressive Trajectory Generation
Changfei Fu, Guangcheng Chen, Aoxiang Gu +3
Benefiting from the powerful priors embedded in large-scale pre-training data and the emerging commonsense reasoning ability, large language models (LLMs) have shown unprecedented…
cs.RO2026
VLAConf: Calibrated Task-Success Confidence for Vision-Language-Action Models
Dehao Huang, Aoxiang Gu, Chengjie Zhang +5
Task-success confidence estimation for Vision-Language-Action (VLA) models provides a crucial task-level signal for monitoring manipulation in open-world environments and supportin…
cs.RO2026
Easy-IIL: Reducing Human Operational Burden in Interactive Imitation Learning via Assistant Experts
Chengjie Zhang, Chao Tang, Wenlong Dong +3
Interactive Imitation Learning (IIL) typically relies on extensive human involvement for both offline demonstration and online interaction. Prior work primarily focuses on reducing…