2 papers
cs.CV2026
CA-OPD: Confidence-Aware On-Policy Distillation for Structured Visual Prediction
Menghao Li, Linjie Mu, Yin Wang +6
Autoregressive vision language models unify heterogeneous perception tasks but are highly susceptible to compounding errors. On-policy distillation (OPD) bridges the training-infer…
cs.RO2026
BICPO-VLA: Behavior-Identified Continuation Preference Optimization for Smooth Asynchronous Vision-Language-Action Control
Ming Shang, Yuchen Huang, Jiaoyang Chen +8
The request-to-handoff gap has three coupled sources: ambiguity about the behavior intended at request time, physical-state drift accumulated during action generation, and residual…