vision-language models 2benchmark 1counterfactual evaluation 1dense reward learning 1failure synthesis 1inference regularization 1language grounding 1reinforcement learning 1robot control 1robotic manipulation 1
From the 2 of 15 linked papers with an AI index.
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
When Vision Overrides Language: Evaluating and Mitigating Counterfactual Failures in VLAs
Yu Fang, Yuchun Feng, Dong Jing +5
The paper studies how Vision-Language-Action (VLA) models often ignore language instructions by relying on visual shortcuts, introduces a counterfactual benchmark (LIBERO-CF) to ev…
cs.CV2025
ReBot: Scaling Robot Learning with Real-to-Sim-to-Real Robotic Video Synthesis
Yu Fang, Yue Yang, Xinghao Zhu +4
Vision-language-action (VLA) models present a promising paradigm by training policies directly on real robot datasets like Open X-Embodiment. However, the high cost of real-world d…