2 papers
cs.RO2026
Causal Reward World Models: Zero-shot Reward Design for Automated Skill Generation
Yang Yang, Yuchuang Tong, Zhengtao Zhang +6
Automated Reward Design (ARD) aims to replace manual reward engineering in reinforcement learning with language-driven reward function synthesis. However, existing approaches based…
cs.CV2026
VTI-CoT: Visual-Textual Interleaved Chain of Thought for Video Reasoning
Shufan Zhang, Ziyue Lin, Bairun Wang +4
Video reasoning aims to understand complex temporal events and causal relationships within videos. Recently, Chain-of-Thought (CoT) has been introduced to this field to enhance rea…